gpu hardware llama-cpp local llms How Much VRAM Do You Actually Need to Run Local LLMs? September 22, 2026 — Jonathan Mitchell