LaunchAI ModelsJuly 23, 2026
NVIDIA unveils NVFP4 for cheaper LLM inference

NVIDIA introduces NVFP4, a new 4-bit floating point format that reduces LLM inference costs while maintaining accuracy. The format is designed for Blackwell GPUs and offers up to 2x throughput improvement.
1 source
More stories today
- Fields medalist Jacob Tsimerman joins OpenAI
- Cerebras and Flex Expand U.S. AI Supercomputer Manufacturing
- Alphabet's Anthropic stake reaches $124 billion
- Andrew Ng releases OpenWorker, open-source desktop AI coworker
- Echo achieves Fable-level results at 1/3 cost using open-weight models