As of Sep 30, community quantizations and local deployment comparisons of open-weight models including Qwen3.8 Flash Next are actively shared on r/LocalLLaMA.
Timeline · newest first
Sep 30
Unconfirmed
ATX-Swift-1.5-Qwen3.8-27B-Uncensored community quantization sustaining 50-65 t/s
A Reddit post introduces ATX-Swift-1.5-Qwen3.8-27B-Uncensored-MTP, a community quantization of Qwen3.8-27B sustaining 50-65+ tokens per second with i1-Q5_K_M format.