Which is better for local LLMs, AMD RX 7900 XTX or NVIDIA RTX 4080 Super?
It depends on your workload. For local AI inference with Ollama, the RTX 4080 Super offers superior memory bandwidth (736 GB/s vs 960 GB/s effective) and native CUDA/Optimus support, but the RX 7900 XTX has 24GB VRAM vs 16GB, often a bottleneck for larger models. The winner depends on model size: 7900 XTX for 13B+ models, 4080 Super for smaller models or tasks needing fast token generation.