MTPLX V2 boosts MLX inference, achieves 82 TPS on Qwen 3.6 27b on MacBook Pro

MTPLX V2 introduces Turbo Mode with custom quantized-matmul kernels, achieving 82 tokens per second on Qwen 3.6 27b running on a MacBook Pro. The update includes a Swift-based app and upgraded CLI for coding use.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- GPT Images viral again on Reddit
- Suno users rush to generate 4.5+ Pro songs before deadline
- Qwen 3.8 35B A3B requested by users for speed
- Why do most tech subs seem to hate Claude so much?
- Tutorial: Build document intelligence pipeline with deepDoctection