Cosmos3 64B INT4 quants run locally on CUDA/MLX

Community release provides INT4 quantized weights for NVIDIA Cosmos3 (64B) text-to-image and image-to-video models, enabling local inference on CUDA and Apple Silicon via MLX. Includes code and a comparison with Grok.
1 source
More stories today
Tool converts PDFs, videos into Neo4j knowledge graphs
Tom Doerr·1 hour ago
Suno removes v4.5, angering loyal users
A Reddit user reports Suno removed v4.5, their preferred version since release, and that v6, like v5 and v5.5, doesn't deliver the raw results they liked. The user spends about $90/month across three accounts.
r/SunoAI·1 hour agoMatt Wolfe builds AI slop detector
Matt Wolfe attempts to build a tool that detects AI-generated videos from YouTube, TikTok, Instagram, or X links. He finds the task harder than expected.
YouTube·1 hour ago
AI Models by email
Get an email when there's news on AI Models
No news that day, no email.
Seahawks use Copilot in Excel for gameday analysis
Seattle Seahawks analysts use Copilot in Excel to surface data from previous plays and similar game situations within the ~25-second window between plays, helping coaches make faster decisions.
YouTube·1 hour ago
Hugging Face CEO urges open AI research to address risk
Clem Delangue·1 hour agoWhatsApp Business platform runs on single binary
Tom Doerr·1 hour ago
AI alignment progress assessed in new overview
Emad Mostaque·1 hour ago
Browserbase adds Stagehand tools for Deep Agents
LangChain·2 hours ago