Paper page - VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models
…This effectively places it in the performance band of first-tier reasoning systems, matching or exceeding flagship models that are orders of magnitude larger, such as DeepSeek V3.2, GLM-5, and…