Growing compute shortage driven by agentic AI workloads, Apollo analysis warns

Agentic AI systems can consume 100x–1,000x more tokens than traditional chatbot requests. Constraints in GPUs, memory, TSMC capacity, and power infrastructure are emerging simultaneously.
1 source
More stories today
Meta's Muse Image debuts in top 5 on Artificial Analysis Text-to-Image Arena
Muse Image (Meta) entered the Artificial Analysis Text-to-Image Arena at #5 with a score of 1311.9. It also debuted at #4 on the Image Editing Leaderboard, landing on the Pareto frontier for quality vs price.
Artificial Analysis Leaderboard·1 hour ago
H3 Max Director: first continuous real-time frontier video model
fal·1 hour ago
NVIDIA shows how to deploy reasoning models on Jetson edge devices
NVIDIA's blog details deploying and optimizing multi-step reasoning and agentic AI models on Jetson edge devices, addressing the challenge of running large models at the edge. It covers JetPack, Jetson Orin, and Thor platforms.
NVIDIA Developer Blog·2 hours ago

Business by email
Get an email when there's news on Business
No news that day, no email.
90M conversational LLM runs on Sony PSP from 2004
A 90M-parameter conversational LLM now runs on the Sony PSP (2004 hardware) at 0.5–0.6 tokens per second, near the console's practical limit. GitHub project LLMPSP enables the port.
r/LocalLLaMA·2 hours ago
Google Cloud DevEx program sprints target enterprise AI governance
Google Cloud's Gemini Enterprise DevEx program runs sprint testing of end-to-end developer workflows to find and fix friction. This sprint focused on enterprise AI governance, covering agent identity provisioning, registry, gateway binding, and policy enforcement.
Google Developers Blog (AI)·2 hours ago

Altman: 38,000 ChatGPT queries use water of one almond
OpenAI CEO Sam Altman says 38,000 ChatGPT queries use as much water as producing one almond in California, claiming data centers use no more water than an office building. He calls water-usage claims outdated and based on misinformation.
r/artificial·2 hours ago
AWS and NVIDIA build Physical AI model factory with Cosmos 3 on SageMaker HyperPod
AWS blog details a continuous pipeline for Physical AI systems (robots, AVs) using NVIDIA Cosmos 3 on SageMaker HyperPod, covering synthetic data generation and post-training of perception and policy models.
AWS AI Blog·2 hours ago

Artificial Analysis Index criticized as misleading for real-world LLM performance
A Reddit user claims Muse Spark 1.3 is not on par with OPUS or SOL, suggesting the Artificial Analysis Index is easy to game. An OpenTeams engineer criticizes the Intelligence vs. cost plot for using log scale and official API pricing, which can misrepresent open-weights model costs.
r/Singularity·2 hours ago