Paper page - Qwen-Image-VAE-2.0 Technical Report
… Qwen-Image-VAE-2.0 achieves state-of-the-art reconstruction performance, demonstrating exceptional capabilities in both general domains and text-rich scenarios at high compression ratio. …
Tracked topic
Qwen3 is an AI model family developed by Alibaba, released as a set of large language models for natural-language tasks.
… Qwen-Image-VAE-2.0 achieves state-of-the-art reconstruction performance, demonstrating exceptional capabilities in both general domains and text-rich scenarios at high compression ratio. …
Papers arxiv:2606.03746 Qwen-Image-Flash: Beyond Objective Design Published on Jun 2 Submitted by Tianhe Wu on Jun 4 Qwen Authors: Tianhe Wu , , , , , , , , , , , , , , , , , , , , , Abstract Few-step distillation for visual generative models benefits from systematic investigation of training recip… …
… Experiments on Qwen3, Ministral, and GLM models reveal a persistent English--Chinese performance gap. …
… Specifically, RACES improves DeepSeek-R1-Distill-Qwen-14B by an average of 3.1 points from 48.2 to 51.3 and boosts Qwen3-14B performance from 58.8 to 61.1 on six benchmarks, which are unseen during the construction of training environments. …
Papers arxiv:2605.30280 Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments Published on May 28 Submitted by taesiri on May 29 2 Paper of the day Qwen Authors: , , , , , , , Zhixuan Liang , , , , , , , Shuai Bai , , , , Gengze Zhou , , , Abstract A u… …
… On AIME and HMMT mathematical reasoning, KVpop retains 98% of full-attention performance on Qwen3-4B at 75% KV cache compression and 97% at 88% compression, consistently outperforming established eviction baselines. Qwen3-8B shows even stronger results, reaching near-full teacher performance. …
… Generated by Qwen/Qwen2.5-Coder-32B-Instruct While large language models have been dominating the research landscape recently, small language models remain highly relevant across various domains; yet, they receive far less attention. …
… On BEIR , KaLM- Reranker -V1 achieves state-of-the-art performance, on par with strong industrial models such as the Qwen3- Reranker series; on MIRACL , despite not being extensively trained on multilingual data, KaLM- Reranker -V1 still shows excellent reranking performance. …
… With Qwen3-4B as the backbone, our framework achieves the strongest aggregate performance on our benchmarks, outperforming larger proprietary LLMs e.g., GPT, Gemini and fixed-environment training baselines. …
… Empirical results across ten benchmarks e.g., VideoMME, LVBench demonstrate that OmniAgent achieves state-of-the-art performance among open-source models. Notably, on LVBench, our 7B agent outperforms the 10times larger Qwen2.5-VL-72B 50.5% vs. …