If you primarily run single-GPU inference with models up to 13B parameters, like Llama 3 8B or Mistral 7B, the Threadripper 7980X is unequivocally the better choice. Its faster cores and cheaper platform cost translate directly to snappier chat experiences and lower wattage during idle periods, which matters for a local setup. The EPYC 9354P only makes sense if you plan to cluster 4+ GPUs or lean heavily on CPU offloading for 70B+ models, where its 12 memory channels become a decisive advantage, but prepare for a significantly higher upfront investment and power bill.
For 95% of the local AI community, the Threadripper 7980X wins on performance-per-dollar, latency, and practicality. The EPYC line remains relevant for enterprise-scale deployment, not for a single-developer home rig. Unless you are building a dedicated multi-user inference server with massive models, save your money and go Threadripper. If you truly need EPYC, consider the 9654 instead for better clock-vs-core balance.