gfx906
Here are 14 public repositories matching this topic...
Open-source local AI server configs, GFX906 runtime maintenance, reproducible benchmarks, and QC methods for affordable AI research infrastructure.
-
Updated
Jul 1, 2026 - Python
FlashAttention-style custom attention backend for vLLM on AMD MI50/MI60/Radeon VII (gfx906). Downstream fork of mixa3607/ML-gfx906 with replacement HIP kernels and a vllm.general_plugins entry point.
-
Updated
Apr 22, 2026 - Python
AMD Instinct MI50/MI60 (gfx906, HIP/ROCm) port of NInfer - the specialized single-GPU Qwen inference engine. First AMD target in the ecosystem. Audit complete, port in progress.
-
Updated
Sep 3, 2026 - C++
ROCm/Unsloth/bitsandbytes 4-bit lab and VRAM benchmarks for AMD MI50/gfx906 LLM fine-tuning
-
Updated
Jun 3, 2026 - Python
openPangu Flash92 — C++/CUDA & C++/HIP ROCm Engines. Fully resident, Blackwell tensor cores, no Python. Powered by openPangu.
-
Updated
Sep 13, 2026 - HIP
Restore videos with pixelated/mosaic regions rocm_gfx906 (mi50 gpu)
-
Updated
Sep 10, 2026 - Python
Run ComfyUI on AMD Radeon VII (gfx906) via Docker. ROCm 5.7 + PyTorch 2.3.1, SDXL 1024×1024 in ~28s/image. Pinned to ComfyUI v0.3.60 — the last gfx906-compatible build.
-
Updated
May 8, 2026 - Python
Benchmarks and runnable setup for Qwen LLMs on AMD Radeon VII (gfx906) via llama.cpp + ROCm in Docker
-
Updated
Jul 8, 2026 - Shell
Qwen 3.ocho (C++): stochastic trajectory rendering with a decision/commit loop on a native CUDA/HIP/ROCm inference engine (NVIDIA GB10 + 4x AMD MI50)
-
Updated
Sep 13, 2026 - Cuda
Custom C++/HIP inference engine for Qwen3.8-27B on 4x AMD MI50 (gfx906/ROCm). Written from scratch - not llama.cpp, not a wrapper, no Python in the execution path. 103.7 tok/s with chained MTP speculative decoding.
-
Updated
Sep 13, 2026 - HIP
Run llama.cpp on Vega 64 / gfx900 and Radeon VII / MI50 / gfx906 with ROCm 7.2 — including mixed-GPU inference.
-
Updated
Aug 25, 2026 - Shell
Add this topic to your repo
To associate your repository with the gfx906 topic, visit your repo's landing page and select "manage topics."