Local LLM setup on M4 Pro Mac Mini with Qwen3.6 and Gemma-4

A developer details running a local LLM server on an M4 Pro Mac mini with 48 GB RAM, using Qwen3.6-35B-A3B-OptiQ-4bit and Gemma-4-E4B-it-OptiQ-4bit via oMLX. Setup takes about 30 minutes and covers privacy, cost, and AI sovereignty motivations.
1 source
Developers by email
Get an email when there's news on Developers
No news that day, no email.
More stories today
- Data center spending to hit $31.6T by 2050 on AI boom
- OpenAI explores hiding model 'thinking', raising safety concerns
- Emad Mostaque: Frontier models will one-shot at 10k tokens/sec
- LangSmith adds Messages View for agent debugging
- Notebook collection covers 30 LLM agent memory techniques