Skip to content
Hacker News front page

Qwen 3.6 27B is the sweet spot for local development

Piotr Migdał tested Qwen 3.6 27B and calls it the first local model that works as a general intelligence. Running 8-bit quantized on a Macbook Max M5 128GB with llama.cpp and multi-token prediction, it hits 32 tok/s using 42GB RAM. It handled constrained writing, generated a hexagonal minesweeper npm package in one shot, and built a reactive landing page. The post includes full llama.cpp setup commands and recommends against Ollama on ethical grounds.

Why it matters: A first-person experiment with real numbers, not a press release. Qwen 3.6 is already a hot topic, and this piece adds practical local-deployment details. Score capped at 72 because it's a personal review, not an official launch or major product update.

Read the original ↗Export Markdown