| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Download Repo ZIP] [Original HTTPS Page] |
LLM speculative inference server for heterogeneous hardware & consumer GPUs
vLLM for AMD RDNA4 (gfx1201): Radeon AI PRO R9700 & RX 9070 XT — native MXFP4 linear kernel + spec-decode verify fix, carried as rebased branches (fork-carry model)
Add a description, image, and links to the r9700 topic page so that developers can more easily learn about it.
To associate your repository with the r9700 topic, visit your repo's landing page and select "manage topics."
| Back | FazBrowse Home | New Git URL |