LocalLLaMA user builds home server for LLM inference, overclocks GPU for high bandwidth
A user on the r/LocalLLaMA community reported building a home server from an old PC with a GPU upgrade, achieving approximately 30 tokens per second when running Qwen3.8 27B. The same user also claimed to have successfully overclocked a CMP 170hx GPU, reaching a memory bandwidth of 1.89 TB/s.
What every outlet reports
- User built a home server
- Server uses an old PC with a GPU upgrade
- Qwen3.8 27B model was run
- Achieved ~30 tokens per second with Qwen3.8 27B
- CMP 170hx GPU was overclocked
- Overclocking reached 1.89 TB/s memory bandwidth
Figures reported differently
Qwen3.8 27B tokens per second~30
CMP 170hx memory bandwidth1.89 TB/s
All 3 articles
optimist2
pragmatist1