Can It Run?
GPU × AI model fit checker — instant VRAM analysis with measured speed data.
Consumer GPU
Radeon RX 7600 XT → Llama 3.1 8B
Radeon RX 7900 XTX → DeepSeek R1-Distill 32B
Radeon RX 7900 XTX → FLUX.1 dev
noindex
Radeon RX 7900 XTX → Llama 3.1 8B
Radeon RX 7900 XTX → Llama 3.3 70B
Radeon RX 7900 XTX → Qwen 3 32B
Arc A770 16GB → Llama 3.1 8B
noindex
Arc B580 → Llama 3.1 8B
Arc B580 → Qwen 3 8B
noindex
Arc B580 → Stable Diffusion XL 1.0
noindex
GeForce RTX 2080 Ti → Llama 3.1 8B
noindex
GeForce RTX 3060 12GB → FLUX.1 dev
GeForce RTX 3060 12GB → Llama 3.1 8B
GeForce RTX 3060 12GB → Qwen 3 8B
GeForce RTX 3060 12GB → Stable Diffusion XL 1.0
noindex
GeForce RTX 3090 → DeepSeek R1-Distill 32B
GeForce RTX 3090 → FLUX.1 dev
noindex
GeForce RTX 3090 → Gemma 3 27B
GeForce RTX 3090 → Llama 3.1 8B
GeForce RTX 3090 → Llama 3.3 70B
GeForce RTX 3090 → Mistral Small 3.2 24B
noindex
GeForce RTX 3090 → Qwen 3 32B
GeForce RTX 3090 → Stable Diffusion 3.5 Large
noindex
GeForce RTX 3090 → Wan 2.2 (A14B & TI2V-5B)
GeForce RTX 4060 Ti 16GB → FLUX.1 dev
noindex
GeForce RTX 4060 Ti 16GB → Llama 3.1 8B
GeForce RTX 4060 Ti 16GB → Qwen 3 8B
GeForce RTX 4060 Ti 16GB → Stable Diffusion XL 1.0
noindex
GeForce RTX 4080 SUPER → Llama 3.1 8B
GeForce RTX 4080 SUPER → Qwen 3 8B
GeForce RTX 4090 → FLUX.1 dev
noindex
GeForce RTX 4090 → Gemma 3 27B
GeForce RTX 4090 → Llama 3.1 8B
GeForce RTX 4090 → Llama 3.3 70B
GeForce RTX 4090 → Mistral Small 3.2 24B
noindex
GeForce RTX 4090 → Qwen 3 32B
GeForce RTX 4090 → Stable Diffusion 3.5 Large
noindex
GeForce RTX 5060 Ti 16GB → FLUX.1 dev
noindex
GeForce RTX 5060 Ti 16GB → Llama 3.1 8B
noindex
GeForce RTX 5060 Ti 16GB → Qwen 3 8B
GeForce RTX 5060 Ti 16GB → Stable Diffusion XL 1.0
noindex
GeForce RTX 5080 → FLUX.1 dev
noindex
GeForce RTX 5080 → Llama 3.1 8B
GeForce RTX 5080 → Qwen 3 8B
GeForce RTX 5090 → DeepSeek R1-Distill 32B
noindex
GeForce RTX 5090 → FLUX.1 dev
noindex
GeForce RTX 5090 → Gemma 3 27B
noindex
GeForce RTX 5090 → Llama 3.1 8B
GeForce RTX 5090 → Llama 3.3 70B
GeForce RTX 5090 → Mistral Small 3.2 24B
noindex
GeForce RTX 5090 → Qwen 3 32B
noindex
GeForce RTX 5090 → Whisper large-v3
noindex
Apple Silicon
Data Center GPU
H200 SXM → Llama 3.3 70B
NVIDIA A100 80GB SXM → Llama 3.3 70B
noindex
H100 SXM → Llama 3.1 8B
H100 SXM → Llama 3.3 70B
Pro GPU
NVIDIA RTX PRO 6000 Blackwell → DeepSeek R1-Distill 32B
noindex
NVIDIA RTX PRO 6000 Blackwell → DeepSeek V4 & V4-Flash
NVIDIA RTX PRO 6000 Blackwell → Llama 3.3 70B
noindex
NVIDIA RTX PRO 6000 Blackwell → Qwen 3 32B
noindex
NVIDIA RTX PRO 6000 Blackwell → Qwen 3 8B
RTX 6000 Ada → Llama 3.1 8B
RTX 6000 Ada → Llama 3.3 70B
RTX A6000 → Llama 3.1 8B