
Unsloth now lets you fine-tune and run 500+ LLMs locally on AMD Radeon, Instinct, and Ryzen GPUs across Windows/WSL/Linux with claimed 2x speedups and 70% less VRAM — breaking the practical NVIDIA-only barrier for local training, plus an MCP endpoint that lets AI clients start/stop training and export GGUFs.
“Starting today, our AMD collaboration, custom Triton kernels, and math algorithms enables you to train and run 500+ models across AMD's Radeon, Instinct, Ryzen and data center GPUs, up to 2× faster with 70% less VRAM and no accuracy loss.”
“Unsloth Dynamic NVFP4 keeps accuracy-sensitive layers in FP8 or BF16 while running the rest in W4A4.”
Checking sign-in…
Loading comments…