⌘K
← All GPUs

GeForce RTX 5080

NVIDIA · Consumer GPU · 2025

16 GB VRAM
960 GB/s
360W TDP
$999 MSRP
Check Latest Price
$999 MSRP
Check Price on Amazon →

We earn a commission from qualifying purchases · Prices may vary

Performance Benchmarks

Image Generation all sourced rows →

ModelMetricValueSource
SDXL Turbo FP16 images per minute 65 img/min
ComfyUI · 2025-06-01
Independent benchmark TechPowerUp GPU Reviews ↗
Stable Diffusion 1.5 FP16
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); 1.344 s/image
procyon image overall score 4,650 score
UL Procyon AI Image Generation · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Stable Diffusion 1.5 INT8 INT8
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); 0.561 s/image
procyon image overall score 55,683 score
UL Procyon AI Image Generation · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Stable Diffusion XL FP16
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); 8.808 s/image
procyon image overall score 4,257 score
UL Procyon AI Image Generation · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗

LLM Inference all sourced rows →

ModelMetricValueSource
Llama-2-13B
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); output rate 83.653 tok/s
procyon text overall score 4,790 score
UL Procyon AI Text Generation (TensorRT) · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Llama-3-8B
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); output rate 136.177 tok/s
procyon text overall score 4,424 score
UL Procyon AI Text Generation (TensorRT) · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Llama-3.1-8B Q4_K_M
Token generation: single user, batch 1; source measured 5090=213 / 4090=127 on same rig, within 2% of myaihardware baseline
tokens per second 132 tok/s
llama.cpp via Ollama · 2026-08-01
Unverified community LocalAI Master GPU Benchmarks ↗
Llama-3.1-8B Q4_K_M
Prompt processing: ~5,200 tok/s reported (approximate, batch 1)
prompt tokens per second 5,200 tok/s
llama.cpp via Ollama · 2026-08-01
Unverified community LocalAI Master GPU Benchmarks ↗
Mistral-7B
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); output rate 163.598 tok/s
procyon text overall score 4,635 score
UL Procyon AI Text Generation (TensorRT) · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Phi-3.5-mini
Standardized UL Procyon suite; same lab platform (ThreadRipper 7980X, driver 571.86); output rate 209.459 tok/s
procyon text overall score 4,400 score
UL Procyon AI Text Generation (TensorRT) · 2025-01-29
Independent benchmark StorageReview GPU Reviews ↗
Qwen3-8B Q4_K_XL
Token generation: 16K context, batch 1, Ubuntu 24.04, on-hardware lab measurement
tokens per second 94.1 tok/s
llama.cpp llama-bench, CUDA 12.8 · 2025-12-09
Unverified community Hardware Corner LLM GPU Rankings ↗

Full Specifications

Vram gb
16 GB
Vram type
GDDR7
Memory bus
256 bit
Memory bandwidth gbps
960 GB/s
Tdp watts
360 W
Cuda cores
10752
Pcie interface
PCIe 5.0 x16
Architecture
Blackwell

Related GPUs

Guides & Tools

CompareAIHardware.com participates in the Amazon Associates program. As an Amazon Associate, we earn from qualifying purchases.