AI evaluator METR hit by credential theft, $600K in model credits stolen

Threat actors stole an API key from METR, a nonprofit that evaluates AI models, leading to $600,000 in public AI model credits being consumed. The attack involved credential theft and probing of the organization's systems.
1 source
More stories today
Reddit compares Sol 5.6 Ultra and Astra Light outputs
A Reddit user shared side-by-side outputs from Sol 5.6 Ultra and Astra 6 Light in work mode, calling both "insane" and saying they wouldn't need anything above Astra Light for real work.
r/Singularity·1 hour ago
Anthropic resets Claude Code usage limits
Anthropic has reset usage limits for Claude Code, restoring full access for users. The reset was confirmed by multiple users on social media on September 4, 2026.
TestingCatalog News·1 hour ago
OpenAI quietly raises 5-hour rate limits ~50% across plans
Kimmonismus·1 hour ago
Cybersecurity by email
Get an email when there's news on Cybersecurity
No news that day, no email.
Grok analyzes 47-minute AI employee transcript
Robert Scoble·1 hour agoEEBench measures whether AI can design circuit boards
EEBench uses atopile to test AI circuit design, avoiding GUI clicking. OpenAI's GPT-6 Astra demo in KiCad sparked the question. Models know electronics but real-world constraints like capacitor behavior remain challenging.
Hacker News·1 hour ago
Databricks achieves extreme efficiency via specialized GPU kernel generation
Databricks details a method for generating specialized GPU kernels to replace generic ones in production inference, aiming for extreme efficiency. The approach targets diverse workloads that generic kernels handle inefficiently.
Databricks Blog·1 hour ago

Benchmark of 21 Qwen3.8 27B variants on 16GB VRAM
A Reddit user benchmarked 21 Qwen3.8 27B variants on an RTX 5080 with 16GB VRAM, testing on C code. Best overall was bartowski/Qwen3.8-27B-IQ4_XS; some quants underperformed.
r/LocalLLaMA·1 hour agoLangSmith showcases agent observability with customer stories
LangChain·2 hours ago