Cloud infrastructure

From fork() to Fleet: Designing an Agent Sandbox Cloud — Abhishek Bhardwaj, OpenAI

From fork() to Fleet: Designing an Agent Sandbox Cloud — Abhishek Bhardwaj, OpenAI

This talk by Abhishek Bhardwaj from OpenAI details the architectural considerations for building secure and scalable AI agent sandboxes in the cloud. It explores runtime isolation technologies (from basic process execution to containers, GVisor, and microVMs), emphasizing the superior security of hardware-virtualized microVMs. The speaker then highlights the critical need for persistent storage, outlining explicit (copy-on-write snapshots) and always-on (tiered block storage) solutions as the next major unlock for agent capabilities. Finally, it touches on orchestration challenges for fleet-level management, including low-latency sandbox creation and snapshot-driven scheduling.

Building Data Centers for GPU Clouds

Building Data Centers for GPU Clouds

Craig Tavares, COO of Buzz HPC, provides an in-depth look at the complexities of building and scaling GPU cloud infrastructure for AI. He covers the critical role of renewable energy and strategic location, the evolution of data center design to handle extreme power densities, the importance of a strong partnership with NVIDIA, and the rise of sovereign mandates shaping the future of AI cloud services.

Monster prompt, OpenAI’s business play, nano-banana and US Open experimentations

Monster prompt, OpenAI’s business play, nano-banana and US Open experimentations

The panel discusses KPMG's 100-page prompt for its TaxBot, debating the future of prompt engineering versus fine-tuning. They also analyze OpenAI's potential move into selling cloud infrastructure, the impressive capabilities of Google's new image model, Nano-Banana, and new AI-powered fan experiences at the US Open.

Solving AI Video: How Fal.ai is making AI Video Generation Fatser & Easier

Solving AI Video: How Fal.ai is making AI Video Generation Fatser & Easier

Fal co-founder Burkay Gur and head of engineering Batuhan Taskaya discuss their journey building a high-performance generative media cloud. They cover their strategic pivot to media models, core optimization principles born from early GPU scarcity, and the development of a customer-obsessed culture to navigate the fast-paced AI model landscape.