Moondream 3.1 vision-language model released

Moondream 3.1 is a 9B-parameter vision-language model with mixture-of-experts (2B active). It offers state-of-the-art visual reasoning and detection, with native query, detect, point, and caption skills, designed to be fast and cheap to deploy.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Stable Diffusion user tests H3 model with Cheers-style script
- Reddit users share impressive image-to-video AI demos
- Reddit reminds users they can legally seed AI models via torrenting
- MiniMax H3 reverse-engineers paintings into basic forms
- OpenAI DevDay Exchange Seoul applications close Sept 4