Inference at scale requires optimization of compute, storage and networking performance. In this session, speakers discuss how IBM, Supermicro, and Kioxia can provide a highly scalable solution for AI inference. The session also explains why KV Cache and AI Data Platform (AIDP) are critical for improving inference response time and optimizing the design of an AI Factory. And the benefits that the combine solution brings for improving time-to-first-token responses and optimizing the data and compute design of AIDP on Supermicro all-flash storage systems.
Forgot Password
Almost there!
We just sent you a verification email. Please verify your account to gain access to
Supermicro Open Storage Summit 2026. If you don’t think you received an email check your
spam folder.
In order to sign in, enter the email address you used to registered for the event. Once completed, you will receive an email with a verification link. Open the link to automatically sign into the site.
Register for Supermicro Open Storage Summit 2026
Please fill out the information below. You will receive an email with a verification link confirming your registration. Click the link to automatically sign into the site.
You are already logged into TheCUBE Network as
You’re almost there!
We just sent you a verification email. Please click the verification button in the email. Once your email address is verified, you will have full access to all event content for Supermicro Open Storage Summit 2026.
I want my badge and interests to be visible to all attendees.
Checking this box will display your presense on the attendees list, view your profile and allow other attendees to contact you via 1-1 chat. Read the Privacy Policy. At any time, you can choose to disable this preference.
Inference at scale requires optimization of compute, storage and networking performance. In this session, speakers discuss how IBM, Supermicro, and Kioxia can provide a highly scalable solution for AI inference. The session also explains why KV Cache and AI Data Platform (AIDP) are critical for improving inference response time and optimizing the design of an AI Factory. And the benefits that the combine solution brings for improving time-to-first-token responses and optimizing the data and compute design of AIDP on Supermicro all-flash storage systems.
William Li
GM, Solution ManagementSupermicro
Anders Graham
Sr. Director, SSD Marketing and Business DevelopmentKIOXIA