Fine-tune LLMs on an AMD GPU on Windows via DirectML — no NVIDIA, no ROCm, no dual boot. Proven with a Mistral-7B LoRA fine-tune through Axolotl on a Radeon RX 7700 XT.
-
Updated
Jul 24, 2026 - Shell
Fine-tune LLMs on an AMD GPU on Windows via DirectML — no NVIDIA, no ROCm, no dual boot. Proven with a Mistral-7B LoRA fine-tune through Axolotl on a Radeon RX 7700 XT.
Usability benchmark for local LLMs on AMD hardware (RDNA4 + Strix Halo) — agentic throughput, self-repair counts and context-depth decay, all parsed from llama.cpp server logs
Run a local LLM on AMD hardware ROCm abandoned. Vulkan + llama.cpp + whisper.cpp on GCN 1st-gen and newer. Tested on a Carrizo iGPU from 2015.
Measured llama.cpp Vulkan tuning for AMD Strix Halo (Ryzen AI MAX+ 395 / gfx1151), plus an eval harness with a catalogue of 13 bugs that each produced a believable wrong number.
To associate your repository with the rocm-alternative topic, visit your repo's landing page and select "manage topics."