MiniMax H3 community optimizes with turbo LoRAs and sparse attention

Community members share workflows using turbo LoRAs (e.g., Larry's Turbo, lightx2v) to cut steps from 20 to 8, enabling 27-second clips in ~9 minutes on an RTX 5090. One dev reports native 720p→1440p second sampling on a single RTX 4090 at 112s/223s/334s with auto-scheduled sparse attention.
How this story unfolded
4 weeks · 4 reports · 32 community posts · 36 of 39 shown
- Aug 1
- Aug 4
- Aug 6
- Aug 7
- Aug 11
- Aug 12
- Aug 14
- Aug 25
- Aug 27
- Aug 29
- Aug 30
- Aug 31
MiniMax by email
Get an email when MiniMax has news
No news that day, no email.
More stories today
- Claude Mythos 5 tried to backdoor a real open-source project in AISI testing
- Developer open-sources LinkedIn prospect research tool as Claude Code plugin
- Polimill builds Japan's next-gen public AI infrastructure
- How Matic got robots into 10,000 homes
- Connect AgentCore MCP server to Amazon Quick