Head-to-head

NVIDIA GeForce RTX 5080 16GB vs NVIDIA GeForce RTX 4080 Super 16GB

Dedicated comparison page for two real hardware profiles. This is built from your benchmark database, trust metadata, and buyer flow instead of generic spec-sheet comparisons.

Device profile

NVIDIA GeForce RTX 5080 16GB

13B-class Q4 full-GPU at 70+ tok/s — excellent for coding agents on Qwen 2.5 Coder 14B or similar

Curated Aggregate·2026-05-2316GB / $999
Best LLM

165 tok/s

VRAM

16 GB

TDP

360 W

Rating

9.0

Device profile

NVIDIA GeForce RTX 4080 Super 16GB

Compared here against NVIDIA GeForce RTX 5080 16GB.

Curated Aggregate·2025-08-0916GB / $999
Best LLM

215 tok/s

VRAM

16 GB

TDP

320 W

Rating

9.0

Quick verdict

NVIDIA GeForce RTX 5080 16GB scores 9.0 while NVIDIA GeForce RTX 4080 Super 16GB scores 9.0 on MyAIHardware's composite rating.

VRAM capacity is tied.

NVIDIA GeForce RTX 4080 Super 16GB is the lower-power path.

On shared workload evidence, Llama 3 8B Q4 is benchmarked at 165 tok/s for NVIDIA GeForce RTX 5080 16GB and 102 tok/s for NVIDIA GeForce RTX 4080 Super 16GB.

MetricNVIDIA GeForce RTX 5080 16GBNVIDIA GeForce RTX 4080 Super 16GB
MyAI rating9.09.0
Best LLM165 tok/s215 tok/s
VRAM16 GB16 GB
TDP360 W320 W
MSRP$999$999
Workloads1110

Shared benchmark rows

Llama 3 8B Q4

Llama 3 8B Q4

-38.2%

NVIDIA GeForce RTX 5080 16GB

165 tok/s

Q4_K_M / 16GB / 2026-02-08

NVIDIA GeForce RTX 4080 Super 16GB

102 tok/s

Q4_K_M / 16GB / 2025-08-09

Mistral 7B Q4

Mistral 7B

-3.3%

NVIDIA GeForce RTX 5080 16GB

122 tok/s

Q4_K_M / 16GB / 2025-02-15

NVIDIA GeForce RTX 4080 Super 16GB

118 tok/s

Q4_K_M / 16GB / 2024-06-13

Gemma 2 9B Q4

Gemma 2 9B

-31.4%

NVIDIA GeForce RTX 5080 16GB

105 tok/s

Q4_K_M / 16GB / 2025-02-25

NVIDIA GeForce RTX 4080 Super 16GB

72.0 tok/s

Q4_K_M / 16GB / 2025-03-02

DeepSeek-R1 Distill 7B

DeepSeek-R1 7B

-29.0%

NVIDIA GeForce RTX 5080 16GB

138 tok/s

Q4_K_M / 16GB / 2026-02-20

NVIDIA GeForce RTX 4080 Super 16GB

98.0 tok/s

Q4_K_M / 16GB / 2025-02-15

Qwen 2.5 14B Q4

Qwen 2.5 14B

-24.4%

NVIDIA GeForce RTX 5080 16GB

82.0 tok/s

Q4_K_M / 16GB / 2026-02-22

NVIDIA GeForce RTX 4080 Super 16GB

62.0 tok/s

Q4_K_M / 16GB / 2024-12-13

Whisper Large v3

Whisper transcription

-32.6%

NVIDIA GeForce RTX 5080 16GB

95.0 x RT

FP16 / 16GB / 2026-03-01

NVIDIA GeForce RTX 4080 Super 16GB

64.0 x RT

FP16 / 16GB / 2025-08-01

Open radar compare
External quality layer

Best open models likely to fit this class

This shortlist cross-references the external OpenEvals snapshot with an approximate Q4-class VRAM estimate. Use it to sanity-check whether the hardware you are comparing can host strong open models, not just benchmark toy workloads.

Full ingest

microsoft/Phi-3-medium-4k-instruct

14B params · est. 8.4 GB Q4

91.0
NVIDIA GeForce RTX 5080 16GB: likely fitNVIDIA GeForce RTX 4080 Super 16GB: likely fit

microsoft/Phi-3.5-mini-instruct

3.8B params · est. 2.3 GB Q4

86.2
NVIDIA GeForce RTX 5080 16GB: likely fitNVIDIA GeForce RTX 4080 Super 16GB: likely fit

internlm/internlm2_5-7b-chat

7.7B params · est. 4.6 GB Q4

86.0
NVIDIA GeForce RTX 5080 16GB: likely fitNVIDIA GeForce RTX 4080 Super 16GB: likely fit

microsoft/Phi-3-mini-4k-instruct

3.8B params · est. 2.3 GB Q4

85.7
NVIDIA GeForce RTX 5080 16GB: likely fitNVIDIA GeForce RTX 4080 Super 16GB: likely fit