AnalysisAI ModelsSeptember 9, 2026

Magic claims >10x more compute-efficient pretraining recipe

Magic says its pretraining recipe matches DeepSeek V4 Pro Base using ~50x fewer FLOPs — roughly half of GPT-3's pretraining compute, or ~$0.5M on GB200. Scaling 10x (~$4M) beat all public open base models on perplexity evals.

1 source

More stories today

Open the live feed