rdna4
Here are 61 public repositories matching this topic...
Windows-only version of ComfyUI which uses AMD's official ROCm and PyTorch libraries to get better performance with AMD GPUs. [auto-installation and popular performance enhancing packages like triton * sage-attention * flash-attention * bitsandbytes included ]
-
Updated
Sep 15, 2026 - Python
The intelligent OptiScaler installer Linux gamers needed. Automates FSR4, XeSS & DLSS configuration with GPU-optimized profiles for RDNA3/4, Arc & RTX cards.
-
Updated
Sep 10, 2026 - Shell
Run Gemma-4-31B at full 256K context on a $1,400 AMD RDNA4 GPU (gfx1201): TurboQuant KV cache + HIP-graph-safe Flash-Attention for llama.cpp, fully measured on real hardware.
-
Updated
Jul 23, 2026 - Python
llama.cpp with native AMD RDNA4 (gfx1201) ROCm 7.11 support - 98.97 tok/s AI inference, competitive with RTX 4070 Ti, 32GB VRAM
-
Updated
Jan 3, 2026 - C++
Hardware-focused llama.cpp fork for Windows and dual AMD RDNA4 RX 9070 XT GPUs: ROCm/HIP + Vulkan backends, PyQt6 GUI (RDNA LLM Studio), MTP speculative decoding, FP8 attention, long-context Qwen3.8 benchmarks
-
Updated
Sep 14, 2026 - C++
Fine-tune your own LLM on an AMD Radeon GPU — the easy, tested way. QLoRA via ROCm on Windows/WSL2 & Linux, a worked Gemma-4 example, a reusable live training dashboard, and a smoke test that proves the loss falls.
-
Updated
Jun 20, 2026 - HTML
Evidence-driven reverse engineering of AMD FSR 4.1: DXIL structure, neural weights, dispatch behavior, reproducible tooling and strict claim verification.
-
Updated
Sep 10, 2026 - Python
Qwen3.6-27B-FP8 at 256k context on dual AMD RDNA4 (gfx1201) — working vLLM recipe, audited measurement harness, and the full debugging record
-
Updated
Sep 8, 2026 - Python
Add this topic to your repo
To associate your repository with the rdna4 topic, visit your repo's landing page and select "manage topics."