rdna3
Here are 52 public repositories matching this topic...
Windows-only version of ComfyUI which uses AMD's official ROCm and PyTorch libraries to get better performance with AMD GPUs. [auto-installation and popular performance enhancing packages like triton * sage-attention * flash-attention * bitsandbytes included ]
-
Updated
Sep 7, 2026 - Python
The intelligent OptiScaler installer Linux gamers needed. Automates FSR4, XeSS & DLSS configuration with GPU-optimized profiles for RDNA3/4, Arc & RTX cards.
-
Updated
Aug 4, 2026 - Shell
TRELLIS (Microsoft's Image-to-3D generator) running on AMD GPUs with ROCm. Includes Gaussian splatting, mesh extraction, and GLB export. Tested on RX 7800 XT.
-
Updated
May 9, 2026 - Jupyter Notebook
A turnkey, fully-local AI workstation engineered for the AMD Ryzen AI Max+ 395. LLM inference, voice, document parsing, browser automation, agents — all on-device.
-
Updated
Jul 30, 2026 - Python
A honest port of vLLM-ROCm for windows.
-
Updated
Aug 23, 2026 - Python
Reproducible RDNA reference for Meta Muse-Glimmer-30B — adapting MI-series ROCm recipes to Ryzen AI and Radeon with vLLM, llama.cpp, DFlash and auditable benchmarks.
-
Updated
Aug 17, 2026 - Python
Evaluation-backed AMD ROCm port of HunyuanOCR-1.5 with reproducible OmniDocBench v1.6 benchmarks across llama.cpp, vLLM, and Transformers on gfx1100.
-
Updated
Sep 7, 2026 - Python
Unlock fast, local LLM inference on AMD-powered mini PCs delivering 65-87 t/s for large models without cloud or subscription costs
-
Updated
Sep 6, 2026 - Shell
Multi-GPU vLLM on consumer Radeon (RX 7900 XT/XTX, gfx1100): root cause and fix for the RCCL "operation cannot be performed in the present state" crash, plus 292 benchmark measurements across five model architectures
-
Updated
Sep 7, 2026 - Python
High-performance FP32 GEMV for AMD RDNA 3 (gfx1100), with reproducible performance and numerical analysis.
-
Updated
Aug 31, 2026 - C++
PyTorch built from source for AMD RDNA 3.5 (gfx1150) — Radeon 890M/880M GPU acceleration
-
Updated
Mar 30, 2026 - Shell
Working ROCm 7.2.x + PyTorch environment for RDNA3. Built after fighting pipeline rats and wondering, “Lisa Su, girl… what are they doing down there.”
-
Updated
Jun 19, 2026 - Python
⚡ Zero-dependency Homebrew tap for high-performance LLMs (CachyLLama, ROCmFPX, llama-ai). Optimized for AMD ROCm 7 (Strix Halo, RDNA3/4), Apple Silicon Metal, & Vulkan RADV.
-
Updated
Sep 7, 2026 - Python
Own inference stack on AMD gfx1100: Mojo GPU kernels, self-describing GGUF engine
-
Updated
Sep 6, 2026 - Mojo
Docker infrastructure for AMD Strix Halo (RDNA 3.5 / gfx1151): PyTorch + ROCm base container and a separate Ollama LLM service. Two folders, two Compose files, one Strix Halo box.
-
Updated
Apr 26, 2026 - Shell
Multi-GPU tensor/context parallel diffusion on AMD ROCm — with the patch that makes it actually work.
-
Updated
Apr 19, 2026 - Python
Add this topic to your repo
To associate your repository with the rdna3 topic, visit your repo's landing page and select "manage topics."