05
DAYS
:
08
HOURS
:
32
MINUTES
:
02
SECONDS
Breaking the Context Wall: Storage for Scalable Agentic AI
August 25, 2026 | 5:00 PM - 6:00 PM UTC
Agentic AI at scale will require high-capacity, low-latency context memory to store and retrieve key value, or KV, cache data used in the inference workflow. New storage tiers and architectures have been proposed to solve this problem by balancing the reprocessing of inference queries against the retrieval of previously computed tokens. This new tier, positioned between the local GPU system SSDs and network storage, has been called “context memory” and allows long context tokens to persist in a large-scale storage array. This session with VAST Data, Solidigm and Supermicro describes the implementation, uses and trade-offs of the context memory tier.
Ben Lee
Director, Solution Management Supermicro
Scott Shadley
Director, Leadership Narrative and Evangelist Solidigm
Anat Heilper
Director of AI Architecture VAST Data
person_outline 3343
CW
Chad W.
GK
Gina K.
AP
Amba P.
CK
Cheryl K.
SM
Srinivas M.
TL
Tony L.
WW
Wendell W.