Tencent releases WeMM-Embedding multimodal embedding models

WeMM-Embedding-9B, built on Qwen3.5, accepts text, images, videos, visual documents, and interleaved multimodal inputs, returning a 4,096-dimensional L2-normalized embedding. Audio input is not supported. Versions include 9B, 4B, and 2B.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Nvidia faces investor scrutiny on AI investments
- Lyft builds self-serve AI agent platform with LangGraph and LangSmith
- LangChain: context engineering is key AI skill
- LangChain introduces Structured Tools for complex agent inputs
- LangChain releases multi-vector retriever cookbooks for RAG on tables, text, and images