AI·Coverage

Compute/demo/2026-09-19

LocalLLaMA user builds home server for LLM inference, overclocks GPU for high bandwidth

A user on the r/LocalLLaMA community reported building a home server from an old PC with a GPU upgrade, achieving approximately 30 tokens per second when running Qwen3.8 27B. The same user also claimed to have successfully overclocked a CMP 170hx GPU, reaching a memory bandwidth of 1.89 TB/s.

What every outlet reports

  • User built a home server
  • Server uses an old PC with a GPU upgrade
  • Qwen3.8 27B model was run
  • Achieved ~30 tokens per second with Qwen3.8 27B
  • CMP 170hx GPU was overclocked
  • Overclocking reached 1.89 TB/s memory bandwidth

Figures reported differently

Qwen3.8 27B tokens per second~30
CMP 170hx memory bandwidth1.89 TB/s

All 3 articles

pragmatist1

Related stories