Skip to content
← All topics

#llama-cpp

4 articles

12 min read

Local AI on 8GB of VRAM: this is how I do

My local AI setup AKA My journey of making a frankly unreasonable number of experiments in making 8GB of VRAM behave like more.

local-aiai-tooling
8 min read

Local AI Is Finally Real. It Is Also Weird, Fragile, and Slightly on Fire.

Sparse models, MoE, quantisation, and better local runtimes have changed what is possible on modest hardware. Local AI is no longer only for people running dual RTX 3090 rigs — but it is still very much for developers who can supervise the machine when it starts confidently sawing through the floorboards.

aiai-tooling