ryzen-ai
Here are 53 public repositories matching this topic...
The fastest way to run Qwen3.8-Flash-Next on Strix Halo (gfx1151)
-
Updated
Sep 4, 2026 - Python
Open-source bring-up + verified matmul on the first-gen AMD XDNA1 (Phoenix/Hawk Point) NPU on Linux — the gen FastFlowLM/Lemonade skip. RyzenAI-npu1, mlir-aie/IRON, XRT.
-
Updated
Aug 6, 2026 - Shell
A two-day, non-expert guide to running local LLMs on an AMD Strix Halo (Ryzen AI Max+ 395) box — full ~120GB memory pool, ROCm backend, NPU in parallel via FastFlowLM, one OpenAI-compatible endpoint.
-
Updated
Jul 25, 2026
The fastest way to run Qwen3.8 27B on Strix Halo (gfx1151)
-
Updated
Sep 5, 2026 - Python
vLLM + Qwen3.6-27B (BF16) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). Vision input, 256K context, /v1/responses with separated reasoning, via TheRock ROCm.
-
Updated
Apr 26, 2026 - Python
Xian-VL monorepo - Multilingual Assistant for Gaming Environments 🧙♂️ and other AI translation/research tools. Powered by Lemonade. 🍋
-
Updated
Sep 2, 2026 - Python
AMD ISP4 camera driver module for Ryzen AI laptops
-
Updated
Sep 2, 2026 - Shell
llama.cpp + Qwen3.6-27B (Q8_0 GGUF) OpenAI-compatible inference server on AMD Strix Halo (Ryzen AI Max+ 395, gfx1151). 256K context, ~7.5 t/s decode via TheRock ROCm Docker.
-
Updated
Apr 26, 2026 - Python
Fix internal speakers on the ASUS ProArt PX13 (HN7306, Strix Halo) under Linux: TAS2783 firmware + ALSA UCM Speaker profile + ACP/SoundWire reset for boot and suspend/resume.
-
Updated
Jun 14, 2026 - Shell
ROCmFPX llama.cpp fork for Windows 🏆 — native build, headless OpenAI-compatible server & benchmarks. Tested on AMD Strix Halo (gfx1151), runs on other GPUs too.
-
Updated
Aug 15, 2026 - PowerShell
Stable Diffusion image generation on AMD Ryzen AI NPUs for Linux
-
Updated
Apr 24, 2026 - Python
ComfyUI on AMD Strix Halo (RDNA 3.5 / gfx1151) via Docker. Ubuntu 26.04 LTS + uv-managed Python 3.12 + pinned TheRock ROCm 7.13 wheels. Fixes the silent CPU fallback Debian / Python 3.13 images hit on gfx1151.
-
Updated
Jul 3, 2026 - Python
Run large LLMs locally on AMD Ryzen AI Max+ 395 (Strix Halo, gfx1151) with ROCmFP4 4-bit quantization. Measured benchmarks, build + serving recipes, and 118 ready-to-run GGUF models.
-
Updated
Sep 4, 2026 - Python
Talos-O (Omni): A sovereign, embodied agentic organism forged on AMD Strix Halo. Integrating the Chimera Kernel (Linux 7.0), Zero-Copy Introspection, and the Phronesis Engine. Built from First Principles.
-
Updated
Apr 21, 2026 - Python
A lightweight TUI monitor for AMD Ryzen AI NPUs
-
Updated
Sep 5, 2026 - Python
Reproducible RDNA reference for Meta Muse-Glimmer-30B — adapting MI-series ROCm recipes to Ryzen AI and Radeon with vLLM, llama.cpp, DFlash and auditable benchmarks.
-
Updated
Aug 17, 2026 - Python
Add this topic to your repo
To associate your repository with the ryzen-ai topic, visit your repo's landing page and select "manage topics."