Here are
7 public repositories
matching this topic...
LLM speculative inference server for heterogeneous hardware & consumer GPUs
The fastest vLLM build for dual R9700's. For the latest updates join the Launch80 discord server.
Updated
Sep 15, 2026
Python
vLLM for AMD RDNA4 (gfx1201): Radeon AI PRO R9700 & RX 9070 XT — native MXFP4 linear kernel + spec-decode verify fix, carried as rebased branches (fork-carry model)
Updated
Sep 3, 2026
Python
App to handle one GPU to multiple models/element with queuing and log tracking all with a local hosted web dashboard
Experimental ROCm 10 compatibility build for Qwen3.8-Flash-Next on 4x AMD R9700
Updated
Sep 8, 2026
Dockerfile
Verified dual-R9700 deployment recipe, configs, scripts, and benchmarks for Dyluhn/R9V Qwen3.8 Flash-Next
Updated
Sep 2, 2026
Shell
Minimal AITER LDS and ROCr idle-CPU fixes for Radeon AI PRO R9700 on vLLM 0.29.0 / ROCm 7.2.3
Updated
Sep 13, 2026
Shell
Add this topic to your repo
To associate your repository with the
r9700
topic, visit your repo's landing page and select "manage topics."
Learn more
You can’t perform that action at this time.