Discussion about this post

User's avatar
Pierce Wiederecht's avatar

Let’s not forget AMD execs were also extremely happy when leaving the Vera presentation because they left realizing that Zen6 Epyc had a larger advantage over Vera than they originally suspected. While Vera may still beat Zen6 per core, that’s now up in the air and what’s not up in the air is that Zen6 Epyc will absolutely crush Vera in maximum per socket performance. The only real advantage Vera will have is being packaged into nvl72 racks. As far as being more efficient than x86, the CPUs use so little power compared to the GPUs in rack-scale setups that any efficiency advantage will be negligible in overall power usage. Nvidia is now at the point of CUDA lock-in being so strong that they produce hardware that’s merely competitive instead of truly better and people like Elon Musk lap it up while proclaiming “nvidia is the best”.

c3dtops's avatar

88 cores on a single monolithic compute die -> That looks to me like the biggest compute chiplet/tile made so far for a server class CPU? Makes me thing about the yield they getting out from TSMC fabs.

It helps with core-to-core latency graph (no cliff dive scenario), but that would make the SOC interconnects massive. More metal wiring (at the lowest metal region, M0,M1), wouldn't that hold back max boost clocks?

On this section below:

The difference really between these two setups is that the clustered setup of EPYC has high latency between clusters but within a cluster the latency is low, whereas Vera and Xeon Mesh setup has uniformly average latency; the different configurations are just engineering tradeoffs.

For cloud virtualization workload -> having a more "dis-aggregated" mesh works better (since cores in a rented cluster works in isolation, mostly)

For AI (Training+Inference), since that workload needs more communication across/between cores, bigger interconnect setup works better to bring down latency? Even though a mesh that big has to run at the "Uncore" frequency.

Is that the right way to think about interconnects selection (from a hyper-scaler perspective) when choosing CPU vendors?

4 more comments...

No posts

Ready for more?