x
Thank you! Your submission has been received and we will follow up shortly.
Oops! Something went wrong while submitting the form.
FMS Speaker Session 2026
Events
July 30, 2026

FMS 2026 Speaker Session with Supermicro

AI inference is breaking the old rules of storage.

As context windows grow, agent sessions multiply, and concurrent users scale, KV cache is emerging as the real performance bottleneck for LLM inference. Traditional storage architectures weren't built for this shift. At FMS 2026, Graid Technology is teaming up with Supermicro to unpack how to build AI storage infrastructure that actually scales with today's inference workloads.
‍

📍 Session: AI Storage Made Better, Together: Scaling Approaches to AI Storage Infrastructure

🗓️ August 4 | 8:35 AM

📌 Santa Clara Convention Center, Mission City Ballroom

‍

Speakers:

🎙️ Randy Kreiser, Field CTO, Graid Technology

🎙️ Paul McLeod, Product Director, Storage Systems, Supermicro

‍

We'll walk through scaling approaches across every tier, from single-server to rack-scale to monolithic deployments, and show how dense NVMe GPU platforms paired with Graid Technology's SupremeRAID™ turn SSDs into a high-performance, protected KV cache tier.If storage is becoming your AI infrastructure bottleneck, this is the session to catch. See you at FMS 2026!

Learn More

News & Resources

"You've got a four-lane highway, and to get to your storage, you have to go down to a one-lane bridge." Garrett McKibben joins Solidigm on TechArena's Data Insights podcast to unpack why KV cache overflow is choking AI inference.
Graid Technology CEO Leander Yu and VP of Product Development Alven Yen explain how moving RAID processing onto the GPU breaks the bottleneck traditional controllers cannot solve, delivering 36 million GPU-initiated IOPS for AI infrastructure.
Join Graid Technology at Dell Technologies Forum Taipei 2026 and discover how SupremeRAID™ AE transforms local NVMe storage into a high-performance, scalable, and protected extended KV cache—reducing redundant computation and accelerating time to first token (TTFT).