Local models

Product · 3 updates · last updated Sep 30

Where it stands

As of Sep 30, community quantizations and local deployment comparisons of open-weight models including Qwen3.8 Flash Next are actively shared on r/LocalLLaMA.

Timeline · newest first
  1. Sep 30
    Unconfirmed

    ATX-Swift-1.5-Qwen3.8-27B-Uncensored community quantization sustaining 50-65 t/s

    A Reddit post introduces ATX-Swift-1.5-Qwen3.8-27B-Uncensored-MTP, a community quantization of Qwen3.8-27B sustaining 50-65+ tokens per second with i1-Q5_K_M format.

  2. Sep 30
    Unconfirmed

    M5 Ultra local test: Qwen3.8 Flash Next vs Laguna S 2.1

    A Reddit post compares Qwen3.8 Flash Next and Laguna S 2.1 running on an M5 Ultra in local testing.

    Source: r/LocalLLaMA
  3. Sep 30
    Unconfirmed

    M5 Ultra local test: Qwen3.8 Flash Next vs Laguna S 2.1

    A Reddit post compares Qwen3.8 Flash Next and Laguna S 2.1 running on an M5 Ultra in local testing.

How we label
Official
An institution announces something that has already happened, through its own channel or a founder or executive account.
Statement by …
A named person's views, predictions, plans or comments, including remarks reported by media when the original is unavailable.
Reported by …
Only media or another third party is saying it. We name the outlet.
Unconfirmed
Circulating online (posts, comments, leaks). No company, named person or outlet has confirmed it.

Items older than 72 hours never go in Top stories and show their original date.

DailyCitedRSSAboutPrivacy