mi50
Here are 12 public repositories matching this topic...
Open-source local AI server configs, GFX906 runtime maintenance, reproducible benchmarks, and QC methods for affordable AI research infrastructure.
-
Updated
Jul 1, 2026 - Python
FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point.
-
Updated
Apr 22, 2026 - Python
AMD Instinct MI50/MI60 (gfx906, HIP/ROCm) port of NInfer - the specialized single-GPU Qwen inference engine. First AMD target in the ecosystem. Audit complete, port in progress.
-
Updated
Sep 3, 2026 - C++
ROCm/Unsloth/bitsandbytes 4-bit lab and VRAM benchmarks for AMD MI50/gfx906 LLM fine-tuning
-
Updated
Jun 3, 2026 - Python
Local Paperless-ngx + Ollama integration for AI-based document titles, correspondents, document types, tags, review workflows, and backfill.
-
Updated
Apr 11, 2026 - Python
openPangu Flash92 — C++/CUDA & C++/HIP ROCm Engines. Fully resident, Blackwell tensor cores, no Python. Powered by openPangu.
-
Updated
Sep 13, 2026 - HIP
Qwen 3.ocho (C++): stochastic trajectory rendering with a decision/commit loop on a native CUDA/HIP/ROCm inference engine (NVIDIA GB10 + 4x AMD MI50)
-
Updated
Sep 13, 2026 - Cuda
Run ComfyUI on AMD Radeon VII (gfx906) via Docker. ROCm 5.7 + PyTorch 2.3.1, SDXL 1024×1024 in ~28s/image. Pinned to ComfyUI v0.3.60 — the last gfx906-compatible build.
-
Updated
May 8, 2026 - Python
Custom C++/HIP inference engine for Qwen3.8-27B on 4x AMD MI50 (gfx906/ROCm). Written from scratch - not llama.cpp, not a wrapper, no Python in the execution path. 103.7 tok/s with chained MTP speculative decoding.
-
Updated
Sep 13, 2026 - HIP
Restore videos with pixelated/mosaic regions rocm_gfx906 (mi50 gpu)
-
Updated
Sep 10, 2026 - Python
Add this topic to your repo
To associate your repository with the mi50 topic, visit your repo's landing page and select "manage topics."