I stopped using Qwen and Gemma after finding a local LLM that actually thinks before answering
… What DeepSeek R1 Distill actually is And it confused me at first DeepSeek built its reputation on R1, a reasoning model trained mostly through reinforcement learning instead of the supervised-heavy approach most labs still lean on, and it landed near OpenAI's o1 on math and coding benchmarks, which… …