People keep lining Qwen3.8-27B up with Opus 4.6, the coding frontier from a few months ago.
So, can that class of model sit on a consumer Mac?
Yes, but you still need to build the stack around it:
> an engine that quotes peak RAM
> a harness with a stop
> and tests you own
If you already use Claude Code or Codex, this is a private loop beside that bill.
Qwen3.8-27B fits a Mac as a 17.6 GB file. But MTPLX still peaks at 23.6 GB once the draft-ahead head is on.
A speed taken on a 128 GB Mac with thinking turned off is a different machine, so leave it out of a 24 GB average.
Two ways this goes wrong:
> You stop at the download, and 24 GB of RAM only loads it tight. A 4-bit copy there is still perfect for most tasks. 32 GB is the first real agent. 48 GB can hold a /goal that a machine can check. 16 GB is not this model.
> Or you stop at the first hour. Ollama is a fine look, but after that you still pick an engine, a harness, and a stop.
Local tokens have no invoice, so a runaway does not get expensive, and you still own the merge. Tests are the stop.
We put the pulls, the install, the skill files, and the git-config check on an interactive HTML tutorial.
Which file to pick if you already live in MLX, and what to refuse to train this week, are a longer argument.
Full breakdown below