Compute/shipped/2026-09-29

Qwen 3.8 27B model can be run on AMD GPUs and NPUs with 100K context

The r/LocalLLaMA community reports that the Qwen 3.8 27B model can now be run on AMD hardware, including RX 7800 XT GPUs and AMD NPUs. Community guides detail how to achieve 100K context length on GPUs and note a decode speed of 1 token per second on NPUs using FastFlowLM. This development enables local model use on AMD systems.

3 articles from 1 outlet covered this story. The underlying claim is sourced from a shipped.

What do all outlets agree on?

1 outlet covered “Qwen 3.8 27B model can be run on AMD GPUs and NPUs with 100K context”. All of them report the following:

  • Qwen 3.8 27B model
  • Runs on AMD GPUs
  • Runs on AMD NPUs
  • Supports 100K context
  • Specific GPU: RX 7800 XT with 16 GB
  • Uses FastFlowLM for NPUs
  • Decode speed on NPUs: 1 token per second

Which outlets covered this?

All 3 articles found on this story, grouped by the stance of the piece. Every link goes to the original publisher.

What related stories are there?

Which companies does this involve?

Get the week in AI in one email

What happened, which outlets reported it, and where their coverage differed. One issue a week.

The first issue hasn’t gone out yet. Subscribe and it’s the one you’ll get.

We’ll send the digest and nothing else. One-click unsubscribe. Privacy.