LaunchAI ModelsAugust 18, 2026

Hugging Face releases quantized MOSS-VL models for local use

FP8 and NF4 versions of MOSS-VL-Instruct and MOSS-VL-Realtime run locally in 24GB VRAM, covering image, video, and real-time streaming understanding. The technical report describes an open vision-language model family co-designed for real-time interaction.

2 sources

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed