Qwen

Model · 8 updates · last updated Sep 29

Where it stands

As of Sep 29, Alibaba's Qwen line released Qwen-Audio-3.1-Realtime; community posts claim first Qwen 4 samples, compare Qwen Next 3.8 against Sonnet 5.5, and report Qwen3.8 runs at 95+ tok/s on a single RTX 3090.

Timeline · newest first
  1. Sep 29
    Unconfirmed

    Community releases GSQ-RCO GGUFs for Qwen3.8-Flash-Next with expert-pruned Coder build

    A Reddit release adds GSQ-RCO GGUF quantizations for Qwen3.8-Flash-Next, including a 50% expert-pruned Coder build at about 1.89 bpw.

    Source: r/LocalLLaMA
  2. Sep 29
    Reported by MarkTechPost

    Alibaba Qwen releases Qwen-Audio-3.1-Realtime full-duplex voice model, per MarkTechPost

    MarkTechPost reports Alibaba's Qwen team released Qwen-Audio-3.1-Realtime, a full-duplex voice model trained to reason, call tools and decide when to speak; on a tau-Voice adaptation, task success rises to 82.0% from 78.4%.

    Source: MarkTechPost
  3. Sep 28
    Unconfirmed

    Qwen3-VL 8B on a laptop claimed to beat GPT-5.6 on messy tax-form documents

    A Reddit user reports Qwen3-VL 8B running on a laptop beat GPT-5.6 on tax forms among 137 messy documents, compared against Opus 5.5, Sonnet 5 and GPT-5.6.

  4. Sep 28
    Unconfirmed

    Qwen3.8 27B claimed at 95+ tok/s through 100K generation on a single 3090

    A Reddit user reports 95+ tokens per second through 100K generated tokens for Qwen3.8 27B at 262K context on a single RTX 3090.

    Source: r/LocalLLaMA
  5. Sep 28
    Unconfirmed

    UltraFast Qwen3.8 Flash recipe claims 74 tok/s on one DGX Spark

    A Reddit user shares an UltraFast Qwen3.8 Flash recipe claiming 74 tokens per second, 212 tokens per second aggregate, on a single DGX Spark.

    Source: r/LocalLLaMA
  6. Sep 28
    Unconfirmed

    Qwen Next 3.8 and 3.8-27B compared against Sonnet 5.5 in community testing

    A Reddit post compares Qwen Next 3.8 and its 27B variant against Sonnet 5.5 low and Sonnet 5.5 medium.

    Source: r/LocalLLaMA
  7. Sep 28
    Unconfirmed

    First days running qwen3.8-flash-next on 4x R9700

    A community post describes early results running qwen3.8-flash-next on four R9700 GPUs.

    Source: r/LocalLLaMA
  8. Sep 28
    Unconfirmed

    Community reports say first Qwen 4 samples approach Fable/Opus-level quality

    Community posts claim the first samples of Qwen 4 are already approaching Fable or Opus-level quality.

How we label
Official
An institution announces something that has already happened, through its own channel or a founder or executive account.
Statement by …
A named person's views, predictions, plans or comments, including remarks reported by media when the original is unavailable.
Reported by …
Only media or another third party is saying it. We name the outlet.
Unconfirmed
Circulating online (posts, comments, leaks). No company, named person or outlet has confirmed it.

Items older than 72 hours never go in Top stories and show their original date.

DailyCitedRSSAboutPrivacy