Ditching the cloud for local AI — how I use two mini PCs to process millions of tokens a day and save money on costly API fees
… Setup The whole process uses LM Studio and runs on a mix of quantized models, generally of Qwen3.5 and 3.6. …
Tracked topic