Mistral launches Mistral Large 4 preview, a 1T-parameter open-weight model
Read original source →mistral.ai
Mistral Large 4 ("Le Chonk") is a 1T-parameter natively multimodal model with 49B active parameters, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters. The preview API is live on Mistral Studio; weights drop end of this month.
How this story unfolded
2 days · 14 reports · 27 community posts · 41 of 47 shown
- Oct 6
Introducing Mistral Large 4mistral.ai
Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of Chinawired.com
Mistral debuts Large 4 ‘Le Chonk', a 1-trillion parameter text output model with high benchmarks planned for open weights releaseventurebeat.com
Mistral unveils new AI model it says rivals best open systems from Chinacnbc.com
Mistral’s new 1T model aims to leapfrog closed and open rivalstechcrunch.com
Mistral Large 4 now available on AI Gatewayvercel.com
Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Modelmarktechpost.com
Mistral’s new AI tried to escape its test environment. In three weeks, anyone can download itthenewstack.io
Mistral AI Drops 'Le Chonk': A Massive AI Model Named After a Cat Memedecrypt.co
Mistral Large 4simonwillison.net
Introducing Mistral Large 4: Le chonksimonwillison.net
- Oct 7
- Oct 8
More stories today
Local voice assistant runs Qwen3.5 4B with tool calling, no GPU
A Reddit builder assembled a local voice assistant on Qwen3.5 4B with skills/tool calling, using Voxtral for speech-to-text and Pocket for text-to-speech, running without a dedicated GPU.
r/LocalLLaMA·1 hour ago
Sakana AI paper proposes MASS for recursive self-improvement
David Ha (hardmaru)·1 hour agoReddit user calculates $20 Claude plan yields ~$1,300 in API usage
A Reddit user measured API costs across all session logs after driving a new $20 subscription to near 100% of its 5-hour limit, then extrapolated. The same method put the $200 tier at roughly $8,000 of usage, though the poster calls that earlier figure imprecise.
r/ClaudeAI·2 hours ago
Amazon AGI Lab demos perception agents for visual verification
Emile Baizel and Shruti Arora of Amazon AGI Lab present two perception agent primitives: visual annotation, where a person selects an element, and a browser verification pass that checks whether a coding agent's change actually appeared.
YouTube·2 hours ago
Snyk's Javier Garza demos AI bill of materials workflow
Talk walks through a CLI scan that turns a repository into a searchable inventory of models, datasets and agents, then surfaces it in a visual dashboard. The inventory is framed as the anchor for assessing AI supply-chain risk.
YouTube·2 hours ago
FriendliAI engineer: inference provider choice decides open-weight model speed
FriendliAI founding engineer Yunmo Koo argues open-weight models like GLM 5.2 and MiniMax M3 now compete with the best proprietary models, and shows GLM 5.2 and Claude Opus 4.8 building similar outputs at different speeds.
YouTube·2 hours ago
Karpathy: LLM capability understanding gap is widening
Andrej Karpathy·3 hours agoAuroraIMG-6M: 6M-parameter text-to-image model released
AuroraIMG-6M is a 6M-parameter text-to-image model that produces recognizable images at 64x64 resolution. The model card includes samples, architecture and training details, and a quick start.
r/StableDiffusion·3 hours ago