Prime Intellect benchmarks 18 frontier models on nanoGPT optimizer
Read original source →primeintellect.ai
Prime Intellect conducted 153 autonomous runs across 18 models, with Fable achieving the top result of 52,726. The benchmark evaluated model performance on the nanoGPT optimizer speedrun, with Opus and Kimi K3 following in the rankings.
1 source
More stories today
LeJEPA enables self-supervised pretraining on your own data
Massimiliano Viola·1 hour ago
Reddit users ask when Flux 3 open weights image edit will release
A single r/StableDiffusion thread asks whether Flux 3 open weights image editing has shipped yet, with no release date, version details, or official confirmation in the post.
r/StableDiffusion·1 hour agoReddit user tests whether people can tell high-end AI models apart
A two-month informal experiment with a sample size of 8 asked everyday people to distinguish between high-end AI models. The author describes it as a curiosity test rather than rigorous research.
r/LocalLLaMA·1 hour ago
CLI tool extracts structured knowledge from documents
Tom Doerr·2 hours ago
Boris Cherny: Claude built it from one prompt, no bugs so far
Boris Cherny·2 hours ago
Thariq: AI agent failures trace to imprecise prompts outside expertise
Thariq·2 hours agoDeepSeek shrinks agent memory cache 437x versus DeepSeek-V1
DeepLearning.AI·2 hours ago
Long piece on KV caching in flow models for image generation
Sayak Paul·2 hours ago