01
Readout / On-Device AI
3.5-bit Qwen3.8-Flash-Next on two RTX 3090s beats BF16 on MMLU-Pro
A 3.5-bit GGUF quantization of the 180B Qwen3.8-Flash-Next model runs on 2× RTX 3090s and scores 2.85 points above its BF16 reference on MMLU-Pro.
By Joshua Hrisko · 5 min read
Sep 15, 2026





















































