OpenAI releases 722 math manuscripts from internal frontier model
Read original source →openai.com
OpenAI published 722 manuscripts (719 after three withdrawals) organized into 372 families, including 90 of the top 500 open problems per Proof of Atlas, with Lean proof formalizations on GitHub. Results came from a single model averaging three hours of compute per solution across ~4,000 attempted problems.
People · Terence Tao
How this story unfolded
5 days · 11 reports · 42 community posts · 53 of 64 shown
- Oct 6
- Oct 7
[AINews] Quasi-Riemann-Hypothesis: OpenAI publishes 722 math papers solving 90 of the top 500 open math problems; “the most significant moment” in >100 years of mathematicslatent.space
Complementary remarks from Gary Marcus and Terence Tao on OpenAI’s giant math dropgarymarcus.substack.com
The Mathocalypsescottaaronson.blog
OpenAI Says a Secret AI Model Cracked Hundreds of Open Math Problems in One Prompt—Mathematicians Want Receiptsdecrypt.co
- Oct 8
- Oct 9
- Oct 10
More stories today
Essay proposes pharma-style funding model for AI neolabs
Essay argues neolabs are asked to deliver both a research breakthrough and a venture-scale business, each a 1-in-100 outcome, making combined odds 1 in 10,000. It proposes borrowing pharma's approach to risky R&D so frontier labs, neolabs, and investors all come out ahead.
Hacker News·1 hour ago
Reddit debate: AI coding gains stall in large enterprises
An r/ExperiencedDevs thread argues AI does not raise velocity much outside startups because it shifts the bottleneck to code review. The poster says political, procedural and practical friction in large enterprises absorbs most of the time anyway, citing academic studies.
r/ExperiencedDevs·1 hour agoUnblocked's relational context engine cuts token burn for coding agents
Peter Werry, founding engineer at Unblocked, argues vector search can't answer queries like "which PRs did I merge last week?" while a relational context engine can. The engine pulls in PRs, Slack threads and incident data beyond code search, reducing token burn.
YouTube·2 hours ago
LocalLLaMA users debate whether local models can replace coding plans
A r/LocalLLaMA thread argues local models plus a good harness are about a year from writing 90% of code, with commenters pointing to Qwen3.8 Next Flash as already close.
r/LocalLLaMA·2 hours agoChart shows compute shifting from pretraining to post-training
A Reddit r/Singularity post tracks how compute allocation moved from pretraining to post-training over the past two years. The post carries no article text, figures, or linked source.
r/Singularity·2 hours ago
Nvidia in early talks to acquire or invest in Reflection AI
Nvidia is in early-stage negotiations to acquire or further invest in Reflection AI, a US developer of open-weight models, per the Financial Times. No financial terms were reported.
Bloomberg Technology·2 hours ago

Developer compares Gemma4-31B, Qwen3.8-27B and 6.1-Sol on real coding tasks
A 20-year software engineer ran local models on tasks analogous to his actual work, testing Gemma4-31B, Qwen3.8-27B and 6.1-Sol. He posted the observations to r/LocalLLaMA as informal field notes rather than a benchmark.
r/LocalLLaMA·2 hours agoGLM-5.5 teased as next Zhipu model after GLM-5.3
Kimmonismus·3 hours ago