Intel LLM-Scaler adds Muse Glimmer 30B support

Intel's LLM-Scaler-vLLM beta 0.21.0-b3 adds same-day support for Meta's Muse-Glimmer-30B with FP8 online quantization on Arc Pro B70, plus DFlash for Muse-Glimmer-30B and Qwen3.6-27B. LLM-Scaler-Omni beta 0.2.0-b1 adds MiniMax H3 video generation and Wan Animate 2 support.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- MEES, Minimax H3 experiment
- llama.cpp PR list targets faster CPU inference
- Tutorial: Build ensemble weather forecasts with NVIDIA Earth2Studio
- AI band gets YouTube Official Artist Channel status
- Sony Music, Warner sue Anthropic over alleged IP theft