FitMyLLM

Running LLMs on AMD Radeon Pro Vega II Duo: What You Can Run

The AMD Radeon Pro Vega II Duo packs 32GB of HBM2 with 1020 GB/s of memory bandwidth, making it one of the best consumer GPUs for running local AI models. Here is which models fit, how fast they run, and what quantisation gives the best balance of quality and speed.

Already have it? Find the best model for it. Still choosing? See what a budget buys. Or the full specifications.

Loading the model tables…

Running LLMs on AMD Radeon Pro Vega II Duo: Complete Guide

The AMD Radeon Pro Vega II Duo has 32GB VRAM and 1020 GB/s memory bandwidth, making it one of the best consumer GPUs for running local AI models. This guide covers which LLMs fit, expected tok/s performance, and recommended settings.

See the full AMD Radeon Pro Vega II Duo specs page for detailed specifications, all compatible models, and speed estimates.