Derek Dicker & Vik Malyala | Beyond the GPU
In this interview from Beyond the GPU: The Full-Stack Fight for Enterprise AI 2026, Derek Dicker, corporate vice president of theย ...
Senior VP, MD & President of the EMEA Division, Supermicro
Search every verified Vik Malyala interview, podcast appearance, and on-the-record quote โ each transcript cross-checked by AI and human review to confirm speaker identity. Vik Malyala, Senior Vice President, Managing Director and President of the EMEA Division at Supermicro, has been discussing the company's AI infrastructure portfolio and the importance of balancing CPU and GPU architectures for enterprise AI workloads. In interviews and presentations, Malyala emphasized that AI infrastructure should not be viewed solely as a decision around GPUs, stating that CPUs, accelerators, memory, networking, storage, cooling, software, and rack integration must work as one system. He noted that with the rise of agentic AI, workloads on the CPU can impact 50 to 90 percent of latency, and called it "a criminal waste not to have the best computation" for orchestration. Malyala also highlighted the rapid pace of technology refresh, which he said is now less than a year, while typical data center projects take 12 to 24 months, creating a fundamental challenge where infrastructure can become outdated before deployment. At COMPUTEX 2026, Malyala showcased Supermicro's next-generation platforms powered by NVIDIA, AMD, Intel, and ARM processors, including the Vera Rubin, HGX B300, and AMD Helios systems. He described the AMD Helios as a "beast" and a "massive 7,000 lb machine." Malyala also discussed the company's data center building block strategy, which includes liquid cooling, rack-scale systems, and infrastructure management software. He referenced estimates of $1.5 trillion in IT infrastructure spending, calling it a "big number," and noted that customers are increasingly exploring on-premises deployment after seeing the costs of public cloud services. Malyala added that Supermicro is working with partners such as AMD to co-design systems that can be quickly taken from proof of concept into production.
“Initially the easiest way is to just go to any kind of a public cloud service provider and then take these tokens and start using everyone is happy until they see the bill. And then all of a sudden now what do they do? There is definite benefits so people are trying to figure out how to bring it on prem.”
“We're talking about 40,000 plus cores per rack. I mean when was the last time you heard about that kind of a density in a rack level for CPU. We're not talking about GPUs. I'm talking about CPUs here.”
“They're talking about around 1 and 1/2 trillion dollars of IT infrastructure spend. 1 and 1/2 trillion dollars, which basically means it could be a conservative number, it could be aggressive number because I have seen numbers going three or four trillion also, but fact is that it's a big number that we all are exposed...”
“The technology refresh is less than a year, less than a year. But a typical data center takes 12 to 24 months if you are lucky, assuming that nothing else is blocking it. This is where the problem fundamentally lies. By the time you initiate a data center project and you finish it, the technology changes are changing w...”
“These deployments are actually quite complex. You're talking about billions and billions of dollars going into each and every data center, and the one thing that people want is stuff to be up and running reliably over a period of time.”
“Make in India is one of those very important factors; we are carefully looking at what actually makes sense to make in India, whether it's through us or through our partners so the investments are put in the right place.”
“We have democratized this whole AI adoption, and one way to do that is to make sure that it is getting adopted by as many developers and as many customers as possible.”
“We work with AMD and we also work with several neo clouds such as vulture, tensor wave, cruso and digital ocean to bring these platforms into their hands so that developers will have access to the systems easily.”
“In addition to that, we also have kept some small number of systems in our jump start program โ one can actually sign up for that and get access to this.”
“The whole idea again here is how do we bring this hardware in different shapes and sizes and capacities for our customers so they can take care of different types of workloads, whether it's training or inferencing or edge applications.”
“This is what actually excites us in working with AMD and the ecosystem to make sure that we have an end-to-end solution for that.”
“Game is still the same, getting better โ people still like to build and experiment. We're seeing repatriation from cloud because of cost, data sovereignty reasons, and the desire to have the flexibility to choose whichever platform or technology they want to adopt.”
“We are the factory that is enabling the AI factories โ you need to bring the infrastructure together in a certain way very quickly and more efficiently and be able to roll in these systems along with the software with NVIDIA and whatnot to get to the customers.”
“It's really a fundamental shift in business โ collaboration over competition. We've incorporated open standards like OpenBMC and OCP NIC/ORv3 GPU accelerator modules to give customers the choice to pick the best designs and platforms for their needs.”
“On an accelerated computer, the most expensive thing in the cluster is GPUs โ we want to keep them busy as much as possible, which means the network must have the lowest latency and storage must be highly performant; we work with many ISVs like VAST Vector, DDN, Qumulo and OSNexus to achieve that.”
In this interview from Beyond the GPU: The Full-Stack Fight for Enterprise AI 2026, Derek Dicker, corporate vice president of theย ...
Vik Malyala, Supermicro | AMD Advancing AI 2026 00:00 - Intro 00:02 - Navigating AI Leadership and Industry Transformation 04:08 - AMD's Role in AI and Partnering Strategies 06:41 - Demand for GPUs and Broad AI Adoption 08:43 - Configurability and Scalability in AI Systems 11:57 - Co-design and Engineering with AMD 15:41 - Optimizing Helios: Balancing Demand and Overcoming Data Center Limits 18:19 - Strategies for Sustainable Market Growth and Innovative User-Centric Design
While surging AI demand is raising Data Center complexity, Supermicro is introducing Data Center Building Block Solutionsยฎ andย ...
Supermicro unveils its latest AI infrastructure portfolio, featuring next generation platforms powered by NVIDIA, AMD, Intel, and ARM. In this exclusive interview with Vik Malyala, Chief Business Officer, President & Managing Director EMEA, Technology & Solutions, Supermicro, discover how the company is advancing AI data centers with liquid cooling, rack scale systems, AI workstations, storage, networking, and infrastructure management. Key Highlights ๐จ Supermicro showcases AI platforms powered by NVIDIA, AMD, Intel, and ARM processors. ๐ฉ Vera Rubin, HGX B300, AMD MI350, MI355, and upcoming Hโฆ
Bring AI to customers at scale, from small enterprises to large data centersโ Vik Malyala, Supermicro EMEA.
Vik Malyala, Senior Vice President at Supermicro, shares how democratizing AI starts with getting hardware and platforms into theย ...
In this interview from SC25, Vik Malyala from Supermicro joins theCUBE's John Furrier and Jackie McGuire to discuss theย ...
At OCP Global Summit 2025, Vik Malyala, SVP of Technology and AI at Supermicro explores some of Supermicro's OCP-inspiredย ...
Intervista a Vik Malyala, President & Managing Director EMEA, SVP Technology & AI di Supermicro durante l'evento AI, HPCย ...
Most people think building AI infrastructure is all about software. But the real bottlenecks? Power, cooling, and optimizingย ...
Sign in to search the full transcript archive, filter by topic, and access every quote from Vik Malyala.
The summary and quote tags on this profile are produced with AI assistance from verified, first-person interview transcripts, then checked by our team to confirm the speaker's identity and the accuracy of every quote. See how we verify →