Trying to self-host LLMs made me realize local AI has a friction problem, not a quality problem
…The first two models I downloaded turned out to be too big for my GPU's 12GB VRAM , meaning I had to undo almost an hour of downloading and once again find…