Stealthium Targets Security Blind Spots in AI Accelerators and Neo-Clouds
Read original source →securityweek.com
Startup Stealthium deploys an in-customer agent to detect attacks invisible to CPU-centric security tools inside AI accelerators and neo-clouds such as CoreWeave and Nebius. The piece warns stealthy neo-cloud compromises could create invisible supply-chain threats, citing the Januscape malware as a recent example.
1 source
More stories today
llama.cpp adds int8 coopmat1 matmul for AMD RDNA3/RDNA4
The Vulkan backend commit (70c4e15) adds an int8 cooperative-matrix matmul path for AMD RDNA3 and RDNA4 GPUs. A user reports a massive performance improvement on a 7900XTX.
r/LocalLLaMA·1 hour ago
Fastino releases GLiNER2.5-Decide, a 340M open-weight decision model that runs on CPU
The 340M-parameter model takes text plus a schema of typed questions and returns structured answers, each with a probability distribution, confidence score, and constraint-feasibility metadata. It is open-weight and runs on CPU.
MarkTechPost·1 hour ago

OpenAI reportedly developing $500 ChatGPT Pro Max tier
TestingCatalog reports the plan would add "Fastest Work and Codex" to ChatGPT Pro, with possibly higher usage limits, and may run on Cerebras infrastructure. OpenAI's pricing page lists only Free, Go, Plus, Pro, Business and Enterprise — no Pro Max.
r/OpenAI·1 hour ago
Reddit user says ChatGPT passes 'Tooth Fairy Test' image prompt
A r/ChatGPT user reports ChatGPT is the first AI to render their prompt for a photorealistic, unsettling Tooth Fairy with rows of collected teeth in her mouth. No other models or comparisons are named.
r/ChatGPT·1 hour ago
Claude Code pipeline produces 3:12 film for $16.87 in OpenRouter spend
A Reddit user orchestrated five AI agents through Claude Code to make "The Most Boring Movie Ever (Ever) Made," a 3:12 video with outtakes about a sock puppet named Gerald facing a beige wall for 7 hours 46 minutes. Total OpenRouter cost: $16.87.
r/ClaudeAI·1 hour ago
Mac Studio M5 Ultra 96GB vs M5 Max 128GB debated for local LLMs
M5 Ultra (30/64) with 96GB offers 1.2 TB/s bandwidth and roughly 1.7x faster generation plus faster prefill; M5 Max (40-core GPU) with 128GB runs 614 GB/s but adds 32GB memory and costs less.
r/LocalLLaMA·1 hour agoagent-shell 0.78 adds persistent prompt and turn steering
The Emacs ACP client adds a writeable prompt that stays available while the agent works, queueing new submissions automatically, plus turn steering for in-flight prompts via the session/steering ACP extension. Two agents join: Antigravity and Qoder.
Lobsters·1 hour ago
Claude Opus generates a punk-style zine about "Claudishness"
Ethan Mollick·2 hours ago