
Agents
Run OpenClaw or Hermes on a local model: which fits your GPU
Ollama's default `gemma4` is the small E4B and `qwen3.5` is the 9B, not what most people assume. What each tag really downloads, what the bigger models need at the 64K context agents use, and what to pick for 12, 16 and 24GB cards.
Lena Fischer · 1h ago · 7 min

