Search

Showing top 115 results for "Qwen 3.8 performance" · from 118 indexed matches

Related topics: Qwen3 Alibaba

Tracked topic

Qwen3

Qwen3 is an AI model family developed by Alibaba, released as a set of large language models for natural-language tasks.

Live now Local one-shots, you say? Yep, we're there with Qwen 3.8 27B
2 live sources Latest signal 2h ago See topic hub

Top stories

Discussions and forums

r/LocalLLaMA · u/Traditional_Bell8153 · 2w ago

Tesla V100 Qwen3.6 27B Performance

Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi coding agent llama.cpp model preset: [*] spec-default = 1 ctx-size = 131072…

r/LocalLLaMA · u/davidthesong · 2w ago

Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash

Qwen3.8-Max (2.4T) is another massive contribution to the open weight community. On benchmarks, it performs closely to Kimi K3 and DeepSeek V4 flash across all categories and is better at coding and software tasks. Qwen3…

r/LocalLLaMA · u/ForsookComparison · 18h ago

I'm really hoping we're in 2026's 2-month-gap between QwQ and Qwen3 right now

QwQ was genuine next-gen performance usable on local hardware, but the massive required context (it's reasoning style was akin to "if I say every possible word, I'll notice the right one!") kinda made it unusable for age…

r/LocalLLaMA · u/Forsaken_Goal3692 · 1d ago

Ultrafast Qwen3-TTS at 34 ms Time-to-First-Audio, Handling 10 Requests Per Second [OSS]

Hey locallama! We recently open sourced a Qwen3-TTS 1.7B implementation that achieves 10 requests per second (RPS) and sub-50 ms p95 time-to-first-audio (TTFA) while maintaining real-time playback on 1 x H100. This exten…

r/LocalLLaMA · u/beneath_steel_sky · 3w ago

Medical model: Reasoning-Medical-27B (Qwen3.6-27B finetune)

From the description: "Reasoning-Medical-27B is designed for universal advanced medical reasoning in professional medicine, medical genetics, college biology/medicine, and clinical knowledge. The model was fine-tuned on …