Krea 2 Turbo quantized to INT4 ConvRot at 11.88 GB

A Reddit user converts Krea 2 Turbo BF16 model to INT4 ConvRot, achieving 11.88 GB with aggressive profile quantizing 173 layers. The INT4 version is faster than existing INT8 ConvRot.
1 source
AI Models by email
Get an email when there's news on AI Models
No news that day, no email.
More stories today
- Claude Code 2.1.258 fixes macOS 12 launch regression
- AfterQuery reportedly becomes Y Combinator's fastest-ever unicorn at $3.2B
- Ethan Mollick: agents now capable of long-running self-organized work
- FDA builds AI-ready data foundation on Databricks for Government
- AWS and Amazon MGM Studios launch gen AI film fund