Multi-Terabyte GPUs the Only Answer for AI Scaling



Are multi-terabyte GPUs and massive frontier models really the answer to enterprise AI, or are we hitting a hard wall of soaring memory costs and power constraints? In this episode of the Tech Field Day Podcast, host Alastair Cooke sits down with technical experts Ron Pagani Jr., Andy Banta, and Brian Martin to examine whether brute-force hardware scaling remains sustainable. The panel debates scale-up vs. scale-out architectures, emphasizing how smart software optimizations—such as persistent KV caching (MinIO MemKV, Solidigm) and intelligent data layers (CTERA)—can eliminate redundant GPU computation and lower costs. They also explore agentic microservices, power and cooling bottlenecks, and why right-sizing AI models for production is critical for achieving real-world enterprise ROI.

Panelists

Alastair is a Tech Field Day event lead at the Futurum group, specializing in Cloud, DevOps, and Edge.

Storage Janitor – seasoned technology professional

Vice President of AI and Datacenter Performance at Signal65

Product Management consultancy Founder with focus on technology for enterprise, spanning edge to cloud

Sign up for updates to
Tech Field day events

Thank you for being part of the Tech Field Day community! Our mailing list is a great way to stay up to date on our events and technical content, and we appreciate your signup.

We promise that we’ll never spam you, send ads, or sell your information. This list will only be used to communicate with our community about our events and content. And we’ll limit it to no more than one message per week.

Although we only need your email address, it would be nice if you provided a little more information to help us get to know you better!