
I compared Ollama, vLLM and LM Studio — only one of them survives a second user
at one person typing, all three are basically tied. add a second and the gap turns into a cliff
Lena Fischer · Aug 22 · 7 min
5 articles · newest first

at one person typing, all three are basically tied. add a second and the gap turns into a cliff
Lena Fischer · Aug 22 · 7 min

Llama is the open model that kicked off the whole local-AI movement, and running it yourself is genuinely one command. Here's how, and which size fits your GPU.
Lena Fischer · Aug 3 · 10 min

OpenAI released open-weight models you can run yourself. The 20B fits a 16GB card thanks to clever quantization; the 120B needs a serious rig. Here's how to run gpt-oss locally.
Lena Fischer · Aug 3 · 10 min

Mistral's models are famously efficient — they punch above their size, so you get a lot on modest hardware. Here's how to run Mistral locally in one command, by size.
Lena Fischer · Aug 3 · 10 min

Gemma 3 is one of the best open models you can run at home, and it's multimodal — it reads images too. Here's how to run it in one command, and which size fits your GPU.
Lena Fischer · Aug 1 · 10 min