m
mnlprogetti

Alex M.

@mnlprogetti

Ollama and Local LLM Setup on Low VRAM Hardware

Italy
English
About me
I run local LLMs on hardware most guides ignore: a 2014 GPU with 2GB of VRAM. Instead of guessing, I measure. I've found up to +278% speed just by tuning one parameter Ollama sets too conservatively by default, and a 2.38x difference between quantizations that most people never test. I help people get Ollama and local LLMs running well on the hardware they already have, whether that's an old GPU, limited VRAM, or a home lab box. You get the exact numbers I measured on your setup, not generic advice, so you know why something works, not just what to type.... Read more

Skills

m
mnlprogetti
Alex M.
Offline • 

See my services

AI Technology Consulting
I will set up and optimize a local llm with ollama on old or low vram hardware