I've been running some of the biggest open-weight LLMs for free on Nvidia's cloud
…These are actually usable models that can power real work, but they're also the kind of models that need a serious server to host. There are some limitations compared to running…