Qwen 3.8 27B model can be run on AMD GPUs and NPUs with 100K context
The r/LocalLLaMA community reports that the Qwen 3.8 27B model can now be run on AMD hardware, including RX 7800 XT GPUs and AMD NPUs. Community guides detail how to achieve 100K context length on GPUs and note a decode speed of 1 token per second on NPUs using FastFlowLM. This development enables local model use on AMD systems.
3 articles from 1 outlet covered this story. The underlying claim is sourced from a shipped.
What do all outlets agree on?
1 outlet covered “Qwen 3.8 27B model can be run on AMD GPUs and NPUs with 100K context”. All of them report the following:
- Qwen 3.8 27B model
- Runs on AMD GPUs
- Runs on AMD NPUs
- Supports 100K context
- Specific GPU: RX 7800 XT with 16 GB
- Uses FastFlowLM for NPUs
- Decode speed on NPUs: 1 token per second
Which outlets covered this?
All 3 articles found on this story, grouped by the stance of the piece. Every link goes to the original publisher.