Meta releases Muse Spark 1.3, claiming frontier coding at 125x lower cost
Meta Superintelligence Labs' fourth Muse Spark release in five months uses ~20% fewer tool calls and ~25% fewer tokens than Muse Spark 1.2. Jiaxuan You cites $0.2 vs $25 per 1M output tokens versus Claude Opus 5; LuminaBench scores it 62 on Artificial Analysis vs 59 for Gemini 3.8 Flash.
People · Mark Zuckerberg
How this story unfolded
2 days · 6 reports · 26 community posts · from Sep 2
- Sep 2
- Sep 3
[AINews] Muse Spark 1.3 matches GPT-5.6-Sol, confirming Meta Superintelligence as the newest Frontier Lab, >90% discount for traininglatent.space
“Google was ahead only a few hours”: Muse Spark 1.3 edges out Gemini as Meta claims its biggest coding leap yetthenewstack.io
Meta says Muse Spark 1.3 has frontier performance — but its best results come from a model developers can’t broadly use yetventurebeat.com
Meta AI Released Muse Spark 1.3: An Agentic Coding Model That Uses ~20% Fewer Tool Calls and ~25% Fewer Tokens Than Muse Spark 1.2marktechpost.com
- Sep 4
- Sep 5
More stories today
ComfyUI user seeks uncensored prompt helper for MiniMax H3 image-to-video
A ComfyUI user reports that an existing prompt enhancer for MiniMax H3 broke image-to-video continuity: the subject appeared in a completely different setting after roughly two seconds of a near-frozen start image.
r/ComfyUI·1 hour agoMystery AI Hype Theater 3000 hosts Te Hiku Media's Keoni Mahelona
Emily M. Bender·1 hour agoByteDance Seed's HarnessDev: Only 34 of 64 Harness Changes Generalize
HarnessDev tests whether LLMs can build and evolve their own agent harnesses, finding self-built harnesses transfer poorly across models. GPT-5 solves 35.2% of Terminal-Bench 2.1 tasks in Terminus 2 but 49.6% in Codex CLI with identical weights.
MarkTechPost·1 hour ago

Podcast episode covers pushback against AI surveillance ed tech
Emily M. Bender·2 hours agoReddit thread asks how much developers rely on Claude for coding
r/ClaudeAI discussion asks developers whether they use Claude mainly for debugging and understanding code or for building complete features and projects, and how much generated code they trust without reviewing it.
r/ClaudeAI·2 hours agoEdward Hughes argues AI lacks scientific taste
Inherent co-founder and Chief Scientist Edward Hughes tells Machine Learning Street Talk that creativity is not optimisation, and that AI's missing capability is choosing which questions are worth asking. He argues scientific judgement must be learned through practice rather than specified.
YouTube·2 hours ago
ACL implements new sustainable reviewing policy
Anna Rogers·2 hours agoCohere talk covers predictive representations for continual RL
Cohere-hosted talk by Raymond Chua addresses continual learning in deep reinforcement learning, framing it as a major unsolved challenge for AI agents. The work centers on predictive representations and memory to let agents adapt in complex, dynamic environments.
YouTube·2 hours ago