Vorch-Streamer: 14B diffusion model for real-time infinite-length avatars

Vorch-Streamer, a 14B-parameter diffusion model, generates real-time audio-driven avatars with unbounded length. The arXiv paper adapts a pretrained bidirectional model to causal, continuous synthesis, tackling audiovisual synchronization and visual consistency for long-form streaming.
2 sources
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Sequoia Capital invests in AI-native video platform Preview
- US Launches Effort to Speed Trade in AI Goods Between Allies
- DeepMind launches SL2T sign language-to-text model
- Liquid AI releases LFM2.5-VL-3B vision-language model for edge
- Grok and Meta's release discussed on ETN podcast episode