Skip to content
#

mi50

Here are 12 public repositories matching this topic...

Custom C++/HIP inference engine for Qwen3.8-27B on 4x AMD MI50 (gfx906/ROCm). Written from scratch - not llama.cpp, not a wrapper, no Python in the execution path. 103.7 tok/s with chained MTP speculative decoding.

  • Updated Sep 13, 2026
  • HIP

Add this topic to your repo

To associate your repository with the mi50 topic, visit your repo's landing page and select "manage topics."

Learn more