Mistral launches Mistral Large 4 preview, a 1T-parameter open-weight model
Read original source →mistral.ai
Mistral Large 4 ("Le Chonk") is a 1T-parameter natively multimodal model with 49B active parameters, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters. The preview API is live on Mistral Studio; weights drop end of this month.
How this story unfolded
2 days · 14 reports · 27 community posts · 41 of 47 shown
- Oct 6
Introducing Mistral Large 4mistral.ai
Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of Chinawired.com
Mistral debuts Large 4 ‘Le Chonk', a 1-trillion parameter text output model with high benchmarks planned for open weights releaseventurebeat.com
Mistral unveils new AI model it says rivals best open systems from Chinacnbc.com
Mistral’s new 1T model aims to leapfrog closed and open rivalstechcrunch.com
Mistral Large 4 now available on AI Gatewayvercel.com
Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Modelmarktechpost.com
Mistral’s new AI tried to escape its test environment. In three weeks, anyone can download itthenewstack.io
Mistral AI Drops 'Le Chonk': A Massive AI Model Named After a Cat Memedecrypt.co
Mistral Large 4simonwillison.net
Introducing Mistral Large 4: Le chonksimonwillison.net
- Oct 7
- Oct 8
More stories today
Google DeepMind's Pushmeet Kohli on AI tools for mathematicians
Pushmeet Kohli·2 hours agoHarvey LAB-AA v1.1 adds hallucination check to legal agent benchmark
Artificial Analysis·2 hours ago
BlockRun and Incarna use Bedrock AgentCore payments for agent inference
AWS details pay-per-inference for AI agents, where purchases are small and frequent — sometimes a fraction of a cent each — and happen inside the agent's loop with no person available.
AWS AI Blog·2 hours ago

NVIDIA KGMON team takes second in KDD Cup 2026 Data Agents
NVIDIA's KGMON team placed second in the KDD Cup 2026 Data Agents competition, building its system around a smaller, more verifiable agent harness rather than a larger model. Teams had to work with a small, fixed LLM, making the harness the main optimization surface.
NVIDIA Developer Blog·2 hours ago

Vibe Data Modeling open-source agent builds and validates data models
Databricks·2 hours ago
OpenAI's annualized revenue nears $50B, $20B below prior $70B figure
OpenAI told investors its annualized revenue is approaching $50 billion, per the FT, roughly $20 billion less than the previously reported $70 billion. The higher figure came from investors' attempts to compare directly with Anthropic's annualized revenue, which counts cloud-partner sales that OpenAI does not.
TechCrunch·2 hours ago

Baikal LoopSR x2 anime upscaler adapts Looped-DiT concept
A community developer built Baikal LoopSR x2, a tiny anime/illustration upscaler that applies the recurrent shared-block concept from Looped-DiT to direct RGB restoration. Released for ComfyUI with weights available.
r/StableDiffusion·2 hours ago
Halogen 0.17.2 hits ~45 tps with Qwen 3.8 Flash Next on Strix Halo
Halogen's 0.17.2 update sustains ~45 tokens/sec decode at high context running Qwen 3.8 Flash Next on a 128GB Strix Halo machine. The model is roughly 177B parameters.
r/LocalLLaMA·2 hours ago