Search

Showing top 8 results for "Qwen 3.8 local performance"

Related topics: Qwen3 Alibaba

Tracked topic

Qwen3

Qwen3 is an AI model family developed by Alibaba, released as a set of large language models for natural-language tasks.

Live now Qwen 3.8 27b helped me with something unique that Opus 4 couldn't - Firmware + Software preservation and emulation on an early 2000's ARM based POS system
Live sources: r/LocalLLaMA
1 live sources Latest signal 6h ago See topic hub
cnx-software.com › 2026 › 06 › …

SpacemiT K3 Pico-ITX RISC-V Chassis Kit Review - Part 2: What works, what doesn't in Bianbu OS 4.0 - CNX Software

… So I restarted to give access to compute on the LAN instead, using 0.0.0.0 for the host: 1 jaufranc @ CNXSOFT - spacemitk3picoitx : ~ / LLM $ llama - server - m Qwen3 - 8B - Q4 K M . gguf - t 8 -- host 0.0.0.0 -- port 8080 Performance is about the same at 4.7 tokens per second; it just consumes few… …

Jun 21, 2026 · Jean-Luc Aufranc (CNXSoft)
cnx-software.com › 2026 › 08 › …

GEEKOM IT13 Max Review - Part 3: Ubuntu 26.04 on an Intel Core Ultra 9 185H mini computer - CNX Software

… 5 - 7B - Instruct - int4 - ov -- local - dir ~ / llmmodels / qwen25 - 7b - np ov - npu aey @ IT13 - Max - CNX : ~ $ python - c ' import openvino genai as ov genai pipe = ov genai.LLMPipeline "/home/aey/llmmodels/qwen25-7b-npu", "NPU" result = pipe.generate "What is the capital of Thailand?" , max n… …

Aug 15, 2026 · Jean-Luc Aufranc (CNXSoft)