Fal's H3 Max generates video faster than real time

Fal post-trained MiniMax's open-weight H3 model and optimized its inference stack to run 35x faster than the official endpoint, enabling faster-than-realtime video generation. The result sparked infinite AI-generated livestreams, which got banned on Twitch and Kick, prompting fal to launch its own live video service.
How this story unfolded
4 days · 1 report · 8 community posts · from Aug 27
MiniMax by email
Get an email when MiniMax has news
No news that day, no email.
More stories today
- Data center spending to hit $31.6T by 2050 on AI boom
- OpenAI explores hiding model 'thinking', raising safety concerns
- Emad Mostaque: Frontier models will one-shot at 10k tokens/sec
- LangSmith adds Messages View for agent debugging
- Notebook collection covers 30 LLM agent memory techniques