
AI
Qwen3-Next 80B-A3B locally: an 80B at home, if you can hold it
Qwen3-Next 80B-A3B activates only 3B parameters per token — so it's fast. But 80B of weights still have to live somewhere. The honest VRAM math, the offload trick, and what it really takes.
Lena Fischer · 3h ago · 7 min