Search

Showing top 2 results for "NVFP4"

Filtered by topic: NVFP4 Clear ✕

Tracked topic

NVFP4

Live now A Qwen3.8-27B NVFP4 (with vision) setup is reported to run on one RTX 5090 using a 451K token KV-cache at ~120 tokens/s under a 400W limit.
Live sources: r/LocalLLaMA
1 live sources Latest signal 10h ago See topic hub