Forget Expensive GPUs, This DIY AI Chatbot Cluster Runs On E-Waste
… Of course, that's not the only model used, and much smaller models offer somewhat better performance. Qwen 3 30B apparently runs at just under 10 tokens per second, while Qwen Coder 30B delivers a bit closer to 9 tokens per second. …