RT Anthropic:我们发布了迄今为止最详细的威胁情报报告。报告涵盖了人们如何试图滥用Claude——用于网络攻击,在……
RT Anthropic: We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, in...
Andrej KarpathyAI2026-09-10
检测和应对AI滥用:2026年9月下载报告网络行动阅读更多监控行动阅读更多影响力行动阅读更多常规武器阅读更多生物滥用阅读更多诈骗和欺诈阅读更多非法蒸馏过去八个月里,我们的威胁情报团队识别并干扰了威胁行为者试图利用Claude进行恶意活动的行动。在本报告中,我们分享了这些行动的案例研究,并描述了自2025年3月、8月和11月的上次威胁报告以来,Claude的恶意使用是如何演变的。在每个案例中,我们都干扰了活动,利用所学知识加强了我们的安全措施,并在适当的情况下与当局和行业合作伙伴分享了情报。本报告涵盖了2025年12月至2026年8月期间在七个危害领域内我们干扰的活动:网络行动、影响力行动、监控、诈骗和欺诈、生物滥用、常规武器开发以及蒸馏。使用了Claude Haiku、Sonnet和Opus模型。没有任何滥用案例涉及Claude Fable或Mythos级模型的使用,除了一个非法蒸馏案例。我们在此分享的案例不是典型的滥用,而是我们迄今为止识别的最显著和最新颖的威胁活动的例子。我们发布这项工作,因为我们认为我们有责任披露我们服务的恶意滥用。随着模型能力的不断增强,除非AI开发者和社会的捍卫者采取行动使它们更安全,否则其风险将会增加。本报告涵盖的威胁行为者包括疑似国家资助的团体、受经济动机驱使的犯罪分子、商业间谍软件供应商、国家宣传机构以及政治动机的个人。案例范围从旨在欺诈用户的虚假约会应用网络到旨在识别和监控异见者的监控系统。复杂且持久的威胁行为者……
原文
Detecting and countering misuse of AI: September 2026Download reportCyber operationsRead moreSurveillance operationsRead moreInfluence operationsRead moreConventional weaponsRead moreBiological misuseRead moreScams and fraudRead moreIllicit distillationRead more Over the past eight months, our Threat Intelligence team identified and disrupted operations in which threat actors tried to use Claude for malicious activity. In this report, we share case studies from those operations and describe how malicious use of Claude has evolved since our previous threat reports in March, August, and November 2025. In each case, we disrupted the activity, used what we learned to strengthen our safeguards, and shared intelligence with authorities and industry partners, where appropriate.This report covers activity we disrupted between December 2025 and August 2026 across seven harm areas: cyber operations, influence operations, surveillance, scams and fraud, biological misuse, conventional weapons development, and distillation. Claude Haiku, Sonnet, and Opus models were used. None of the misuse cases involved the use of Claude Fable or Mythos-class models, with the exception of one illicit distillation case.The cases we share here aren’t typical misuse, but rather examples of the most notable and novel threat activity we’ve identified to date. We’re publishing this work because we believe we have a responsibility to disclose malicious misuse of our services. As models become increasingly capable, their risks will increase, unless AI developers and society’s defenders act to make them safer.The threat actors covered in this report include suspected state-sponsored groups, financially motivated criminals, commercial spyware vendors, state propaganda institutions, and politically motivated individuals. The cases range from a network of fake dating apps designed to defraud users to surveillance systems built to identify and monitor dissidents.Sophisticated and persistent threat actors continuously test our safeguards and try to circumvent the technical measures we use to detect and prevent misuse. We’ll continue to evolve our safeguards and coordinate with our partners to improve our ability to detect, disrupt, and prevent future misuse.We hope that the findings in this report will help other developers recognize similar patterns on their own platforms, give governments and civil society a clearer view of how emerging threats take shape, and strengthen collective defenses.