
ikyle.me
June 12, 2026
9 min read
67/100
Summary
Gemma 4 26B-A4B and Qwen3.6 35B-A3B can be run locally on macOS using llama.cpp, MTP speculative decoding, and multimodal support. The setup aims to provide a fast and reliable local coding agent to avoid interruptions from internet failures.
Key Takeaways
Community Sentiment
Positives
Concerns

I ran Gemma 4 as a local model in Codex CLI
Apr 12, 2026

Running Gemma 4 locally with LM Studio's new headless CLI and Claude Code
Apr 5, 2026

Running local models is good now
Jun 16, 2026

Qwen 3.8 27B is excellent, but it defaults to overthinking things
Aug 16, 2026

Qwen 3.6 27B is the sweet spot for local development
Jun 29, 2026