rdna3
Here are 54 public repositories matching this topic...
Windows-only version of ComfyUI which uses AMD's official ROCm and PyTorch libraries to get better performance with AMD GPUs. [auto-installation and popular performance enhancing packages like triton * sage-attention * flash-attention * bitsandbytes included ]
-
Updated
Sep 21, 2026 - Python
The intelligent OptiScaler installer Linux gamers needed. Automates FSR4, XeSS & DLSS configuration with GPU-optimized profiles for RDNA3/4, Arc & RTX cards.
-
Updated
Sep 10, 2026 - Shell
TRELLIS (Microsoft's Image-to-3D generator) running on AMD GPUs with ROCm. Includes Gaussian splatting, mesh extraction, and GLB export. Tested on RX 7800 XT.
-
Updated
May 9, 2026 - Jupyter Notebook
A turnkey, fully-local AI workstation engineered for the AMD Ryzen AI Max+ 395. LLM inference, voice, document parsing, browser automation, agents — all on-device.
-
Updated
Jul 30, 2026 - Python
Multi-GPU tensor-parallel vLLM on AMD Radeon RX 7900 XT / XTX / GRE (7900XT, 7900XTX, RX7900XT, gfx1100, RDNA3, ROCm): root cause and fix for the RCCL hostcall / PCIe atomics (AtomicOps) crash "NCCL error: unhandled cuda error" / "operation cannot be performed in the present state", Proxmox VFIO passthrough; LLM inference benchmarks on 13 machines
-
Updated
Sep 17, 2026 - Python
A honest port of vLLM-ROCm for windows.
-
Updated
Aug 23, 2026 - Python
High-performance FP32 GEMV for AMD RDNA 3 (gfx1100), with reproducible performance and numerical analysis.
-
Updated
Aug 31, 2026 - C++
Reproducible RDNA reference for Meta Muse-Glimmer-30B — adapting MI-series ROCm recipes to Ryzen AI and Radeon with vLLM, llama.cpp, DFlash and auditable benchmarks.
-
Updated
Sep 14, 2026 - Python
AMD/HIP compatibility fixes and GPU-side autoregressive optimizations for Fish Audio S2 Pro on GGML-based C++ runtimes.
-
Updated
Sep 5, 2026 - Cuda
🔥 Automated nightly & release builds of ROCmFPX, Ciru ROCmFPX (DualView Q7/Q8), and q38rocm with AMD ROCm 7 for Windows & Linux. Portable standalone binaries with bundled runtime libraries for Strix Halo (gfx1151), RDNA4, RDNA3, and Homebrew tap.
-
Updated
Sep 21, 2026 - Shell
Unlock fast, local LLM inference on AMD-powered mini PCs delivering 65-87 t/s for large models without cloud or subscription costs
-
Updated
Sep 21, 2026 - Shell
Evaluation-backed AMD ROCm port of HunyuanOCR-1.5 with reproducible OmniDocBench v1.6 benchmarks across llama.cpp, vLLM, and Transformers on gfx1100.
-
Updated
Sep 18, 2026 - Python
PyTorch built from source for AMD RDNA 3.5 (gfx1150) — Radeon 890M/880M GPU acceleration
-
Updated
Mar 30, 2026 - Shell
Experimental Linux/Vulkan video player adapting AMD FSR 4.1-style temporal reconstruction to ordinary decoded video — no frame interpolation.
-
Updated
Sep 16, 2026 - C++
Add this topic to your repo
To associate your repository with the rdna3 topic, visit your repo's landing page and select "manage topics."