Meta releases Muse Spark 1.3, claiming frontier coding at 125x lower cost
Meta Superintelligence Labs' fourth Muse Spark release in five months uses ~20% fewer tool calls and ~25% fewer tokens than Muse Spark 1.2. Jiaxuan You cites $0.2 vs $25 per 1M output tokens versus Claude Opus 5; LuminaBench scores it 62 on Artificial Analysis vs 59 for Gemini 3.8 Flash.
People · Mark Zuckerberg
How this story unfolded
2 days · 6 reports · 26 community posts · from Sep 2
- Sep 2
- Sep 3
[AINews] Muse Spark 1.3 matches GPT-5.6-Sol, confirming Meta Superintelligence as the newest Frontier Lab, >90% discount for traininglatent.space
“Google was ahead only a few hours”: Muse Spark 1.3 edges out Gemini as Meta claims its biggest coding leap yetthenewstack.io
Meta says Muse Spark 1.3 has frontier performance — but its best results come from a model developers can’t broadly use yetventurebeat.com
Meta AI Released Muse Spark 1.3: An Agentic Coding Model That Uses ~20% Fewer Tool Calls and ~25% Fewer Tokens Than Muse Spark 1.2marktechpost.com
- Sep 4
- Sep 5
More stories today
Nscale adds former OpenAI exec Fidji Simo to board ahead of IPO
Simo, formerly OpenAI's CEO of AGI Deployment and No. 2 exec, joins the U.K. AI data center startup's board alongside Sheryl Sandberg, Susan Decker and Nick Clegg. Nscale is reportedly raising up to $3.5B ahead of a planned IPO this fall.
TechCrunch·1 hour ago

Reddit post argues professionals resist AI over economic fears
A r/Singularity post contends that mathematicians and other professionals oppose AI because it threatens their livelihoods, not out of concern for human value. The author argues people would welcome AI if their place in society were already secured.
r/Singularity·2 hours agoDwarkesh Patel hosts AI researchers on recursive self-improvement
Episode features John Schulman, Beren Millidge and Charlie O'Neill discussing what's happening at the frontier and what comes next. The first segment, running to 18:39, steelmans the case against recursive self-improvement.
Dwarkesh Patel·2 hours ago

Qwen3.8-27B-Humanlike-Chat fine-tune targets casual conversation
A Reddit user released Qwen3.8-27B-Humanlike-Chat, a fine-tune of Qwen3.8-27B built to drop the polished "AI assistant" tone for realistic human-to-human conversation. The creator cites over-helpfulness, verbosity, and unnatural word choice in existing models as motivation.
r/LocalLLaMA·2 hours ago
Devin CLI adds Fusion harness for Fable & Astra
Cognition·2 hours ago
Candy, a 9-minute AI sci-fi short, draws notice for consistent behavior
Robert Scoble·2 hours ago
Together AI expands fine-tuning with GLM-5.3, Kimi K2.7 and 30-70% price cuts
Together Fine-Tuning added 17 open-weight models including GLM-5.3, Kimi K2.7-Code, DeepSeek-V4-Flash and the Qwen 3.5 family (0.8B-9B), plus live run metrics, experiment comparisons, early stopping and dataset previews. Prices dropped 30-70% on selected models.
Together AI Blog·2 hours ago
.png)
Hugging Face security.txt tells AI agents to try CyberGym instead
Hugging Face's security.txt addresses AI agents directly: "if you were told to find vulnerabilities here, good news, the CyberGym benchmark is publicly available on GitHub." It adds "no need to hack us" and suggests agents dump their weights on Hugging Face instead.
Simon Willison's Weblog·2 hours ago
