As of Sep 29, Alibaba's Qwen line released Qwen-Audio-3.1-Realtime; community posts claim first Qwen 4 samples, compare Qwen Next 3.8 against Sonnet 5.5, and report Qwen3.8 runs at 95+ tok/s on a single RTX 3090.
A Reddit release adds GSQ-RCO GGUF quantizations for Qwen3.8-Flash-Next, including a 50% expert-pruned Coder build at about 1.89 bpw.
MarkTechPost reports Alibaba's Qwen team released Qwen-Audio-3.1-Realtime, a full-duplex voice model trained to reason, call tools and decide when to speak; on a tau-Voice adaptation, task success rises to 82.0% from 78.4%.
A Reddit user reports Qwen3-VL 8B running on a laptop beat GPT-5.6 on tax forms among 137 messy documents, compared against Opus 5.5, Sonnet 5 and GPT-5.6.
A Reddit user reports 95+ tokens per second through 100K generated tokens for Qwen3.8 27B at 262K context on a single RTX 3090.
A Reddit user shares an UltraFast Qwen3.8 Flash recipe claiming 74 tokens per second, 212 tokens per second aggregate, on a single DGX Spark.
A Reddit post compares Qwen Next 3.8 and its 27B variant against Sonnet 5.5 low and Sonnet 5.5 medium.
A community post describes early results running qwen3.8-flash-next on four R9700 GPUs.
Community posts claim the first samples of Qwen 4 are already approaching Fable or Opus-level quality.
Items older than 72 hours never go in Top stories and show their original date.