Which is better for local LLMs, NVIDIA RTX 5080 or NVIDIA RTX 4090?
It depends on your workload. The RTX 5080 offers competitive raw FP16 throughput and improved memory bandwidth efficiency at a lower price, but the RTX 4090 retains a lead in VRAM capacity (24GB vs 16GB) and mature software support for large local models. For most local AI builders running 7B-13B parameter models, the 5080 delivers better performance-per-dollar; however, for 30B+ models or fine-tuning, the 4090’s extra VRAM is non-negotiable.