The GitHub request to keep MoE experts in RAM is still open after 14 months. But since Ollama 0.30, llama.cpp's own fit step can park them there for…
The short answer
Read this before you set it
Aliteq
Ollama has no --n-cpu-moe switch. Here's what it does with MoE experts instead