The Architecture of "And": Uniting CPU & GPU for Enterprise AI
This theCUBE Power Panel at Advancing AI examines uniting central processing unit, CPU and graphics processing unit, GPU architectures to deliver balanced enterprise artificial intelligence, AI infrastructure. Dave Vellante of SiliconANGLE hosts the session produced with theCUBE Research. Derek Dicker of AMD corporate vice president enterprise business group and Vik Malyala of Supermicro chief business officer discuss engineering collaboration and product roadmaps for Venice sixth-generation EPYC processors, Supermicro H-series servers and the Helios rack-scale architecture for deploying balanced CPU and GPU systems in enterprise environments. Dicker and Malyala discuss Venice performance and Peripheral Component Interconnect Express Gen6 capabilities, Supermicro H-series designs and rack-scale GPU clusters. Dicker explains processor performance characteristics and platform integration; they describe how memory, I/O and networking and cooling integrate to support agentic and inferencing workloads. Malyala outlines system and rack-scale design considerations; they emphasize balanced platforms, liquid cooling and rack-scale planning to maximize cores, memory and I/O. Key takeaways include consolidation and efficiency gains. Dicker notes Turin- and Venice-based systems can replace multiple legacy servers, reduce power use and lower total cost of ownership, TCO while freeing space for AI deployments. Malyala stresses balanced platforms and liquid cooling to maximize performance density. Both recommend rapid proof of concept, POC engagements and close vendor engineering collaboration to accelerate time-to-value and to avoid running agentic workloads on legacy infrastructure.