Cloudflare launches @cloudflare/computer, an agent runtime beyond containers
Read original source →blog.cloudflare.com
@cloudflare/computer, now in early preview, is an agent runtime where the platform decides whether code runs in an isolate, container sandbox, or web browser — giving each agent its own computer. Cloudflare argues container-per-agent setups won't scale to billions of concurrent agents.
1 source
More stories today
llama.cpp adds int8 coopmat1 matmul for AMD RDNA3/RDNA4
The Vulkan backend commit (70c4e15) adds an int8 cooperative-matrix matmul path for AMD RDNA3 and RDNA4 GPUs. A user reports a massive performance improvement on a 7900XTX.
r/LocalLLaMA·1 hour ago
Fastino releases GLiNER2.5-Decide, a 340M open-weight decision model that runs on CPU
The 340M-parameter model takes text plus a schema of typed questions and returns structured answers, each with a probability distribution, confidence score, and constraint-feasibility metadata. It is open-weight and runs on CPU.
MarkTechPost·1 hour ago

OpenAI reportedly developing $500 ChatGPT Pro Max tier
TestingCatalog reports the plan would add "Fastest Work and Codex" to ChatGPT Pro, with possibly higher usage limits, and may run on Cerebras infrastructure. OpenAI's pricing page lists only Free, Go, Plus, Pro, Business and Enterprise — no Pro Max.
r/OpenAI·1 hour ago
Reddit user says ChatGPT passes 'Tooth Fairy Test' image prompt
A r/ChatGPT user reports ChatGPT is the first AI to render their prompt for a photorealistic, unsettling Tooth Fairy with rows of collected teeth in her mouth. No other models or comparisons are named.
r/ChatGPT·1 hour ago
Claude Code pipeline produces 3:12 film for $16.87 in OpenRouter spend
A Reddit user orchestrated five AI agents through Claude Code to make "The Most Boring Movie Ever (Ever) Made," a 3:12 video with outtakes about a sock puppet named Gerald facing a beige wall for 7 hours 46 minutes. Total OpenRouter cost: $16.87.
r/ClaudeAI·1 hour ago
Mac Studio M5 Ultra 96GB vs M5 Max 128GB debated for local LLMs
M5 Ultra (30/64) with 96GB offers 1.2 TB/s bandwidth and roughly 1.7x faster generation plus faster prefill; M5 Max (40-core GPU) with 128GB runs 614 GB/s but adds 32GB memory and costs less.
r/LocalLLaMA·1 hour agoagent-shell 0.78 adds persistent prompt and turn steering
The Emacs ACP client adds a writeable prompt that stays available while the agent works, queueing new submissions automatically, plus turn steering for in-flight prompts via the session/steering ACP extension. Two agents join: Antigravity and Qoder.
Lobsters·1 hour ago
Claude Opus generates a punk-style zine about "Claudishness"
Ethan Mollick·2 hours ago