AIyour GPU isn't too small for that 235B model — you're just using it wrongA handful of llama.cpp flags let a 16GB card run models that shouldn't fit, and most people running local AI have never touched them.Lena Fischer · 1h ago · 7 min