Anthropic threat report details Claude misuse by hackers and states

Anthropic's 154-page report covers misuse of Claude between December 2025 and August 2026, including a Russian state group it calls GTG-20006 that used AI agents to autonomously rebuild malware after detection. It also cites a Yemen-based group using Claude for missile systems and Chinese labs including Alibaba and Moonshot using Claude outputs to train their own models.
How this story unfolded
1 day · 8 reports · 9 community posts · from Sep 10
- Sep 10
- Sep 11
Chinese AI labs secretly used millions of Claude exchanges to train their models, Anthropic sayscnbc.com
Anthropic Says Russian Hackers Used Claude AI to Automate Malware Evasionsecurityweek.com
Anthropic Says Yemen Group Used Claude in Missile Developmentbloomberg.com
Anthropic Says US Adversaries Aimed Claude at Weapons Researchbloomberg.com
Claude users found ways around safeguards for bioweapons researcharstechnica.com
Claude Used to Automate Exploitation and Data Theft Across Multiple Victimsthehackernews.com
Russian State-Sponsored Hackers Use Claude to Rebuild Malware After Detectionthehackernews.com
More stories today
Nscale adds former OpenAI exec Fidji Simo to board ahead of IPO
Simo, formerly OpenAI's CEO of AGI Deployment and No. 2 exec, joins the U.K. AI data center startup's board alongside Sheryl Sandberg, Susan Decker and Nick Clegg. Nscale is reportedly raising up to $3.5B ahead of a planned IPO this fall.
TechCrunch·1 hour ago

Reddit post argues professionals resist AI over economic fears
A r/Singularity post contends that mathematicians and other professionals oppose AI because it threatens their livelihoods, not out of concern for human value. The author argues people would welcome AI if their place in society were already secured.
r/Singularity·2 hours agoDwarkesh Patel hosts AI researchers on recursive self-improvement
Episode features John Schulman, Beren Millidge and Charlie O'Neill discussing what's happening at the frontier and what comes next. The first segment, running to 18:39, steelmans the case against recursive self-improvement.
Dwarkesh Patel·2 hours ago

Qwen3.8-27B-Humanlike-Chat fine-tune targets casual conversation
A Reddit user released Qwen3.8-27B-Humanlike-Chat, a fine-tune of Qwen3.8-27B built to drop the polished "AI assistant" tone for realistic human-to-human conversation. The creator cites over-helpfulness, verbosity, and unnatural word choice in existing models as motivation.
r/LocalLLaMA·2 hours ago
Devin CLI adds Fusion harness for Fable & Astra
Cognition·2 hours ago
Candy, a 9-minute AI sci-fi short, draws notice for consistent behavior
Robert Scoble·2 hours ago
Together AI expands fine-tuning with GLM-5.3, Kimi K2.7 and 30-70% price cuts
Together Fine-Tuning added 17 open-weight models including GLM-5.3, Kimi K2.7-Code, DeepSeek-V4-Flash and the Qwen 3.5 family (0.8B-9B), plus live run metrics, experiment comparisons, early stopping and dataset previews. Prices dropped 30-70% on selected models.
Together AI Blog·2 hours ago
.png)
Hugging Face security.txt tells AI agents to try CyberGym instead
Hugging Face's security.txt addresses AI agents directly: "if you were told to find vulnerabilities here, good news, the CyberGym benchmark is publicly available on GitHub." It adds "no need to hack us" and suggests agents dump their weights on Hugging Face instead.
Simon Willison's Weblog·2 hours ago
