Ai infrastructure

"We're Not Writing Code by Hand Anymore. That's Over." | Owen Jennings & David Haber - The a16z Show

"We're Not Writing Code by Hand Anymore. That's Over." | Owen Jennings & David Haber - The a16z Show

Owen Jennings, Executive Officer at Block, details the company's radical restructuring (40% workforce reduction) driven by AI's impact on productivity. He explains how Block is now operating with smaller squads, leveraging internal AI tools like Goose and Builder Bot, and shipping AI-native products like Money Bot and Manager Bot to deliver personalized, generative UIs for millions of users, emphasizing a future where unique understanding forms the ultimate business moat.

AI Models as a Service: Powering Agentic AI, Privacy, & RAG

AI Models as a Service: Powering Agentic AI, Privacy, & RAG

Cedric Clyburn explains the Models-as-a-Service (MaaS) pattern, detailing how organizations can build their own private AI infrastructure to deploy models like LLMs securely and at scale. He covers the benefits over public APIs, including cost control, data sovereignty, and lifecycle management, and outlines a technical architecture using Kubernetes, API gateways, and observability tools.

One Size Fits None: How Platform Engineering Must Evolve • William Rizzo & Colin Griffin • GOTO 2026

One Size Fits None: How Platform Engineering Must Evolve • William Rizzo & Colin Griffin • GOTO 2026

Colin Griffin and William Rizzo discuss the future of platform engineering, emphasizing the need to move beyond one-size-fits-all frameworks. They explore how different industries like fintech, telco, and automotive require tailored platforms due to unique regulatory and business challenges. The conversation highlights the growing pressure to link platform investments to clear business outcomes and details the infrastructure reckoning caused by large-scale GPU investments for AI, which brings hardware and network orchestration to the forefront.

Greetings, Earthlings: Philip Johnston of Starcloud on Data Centers in Space

Greetings, Earthlings: Philip Johnston of Starcloud on Data Centers in Space

Philip Johnston of Starcloud argues that space will become the primary location for AI compute within a decade. He explains how plummeting launch costs, superior solar energy economics in orbit, and the physics of heat dissipation will soon make space-based data centers cheaper and more scalable than their terrestrial counterparts, predicting a future where nearly a trillion dollars in annual CapEx shifts to space.

The Future of Search: Agents, RAG, and Why Retrieval Still Matters — Simon Eskildsen, Turbopuffer

The Future of Search: Agents, RAG, and Why Retrieval Still Matters — Simon Eskildsen, Turbopuffer

Simon Hørup Eskildsen, founder of turbopuffer, shares his journey from scaling Shopify's infrastructure to creating a new search engine for the AI era. He discusses how a prohibitively expensive experiment at Readwise inspired him to build a cost-effective vector search solution based on object storage and NVMe. Eskildsen breaks down turbopuffer's architecture, its role in cutting costs for companies like Cursor and Notion, his philosophy on building a 'P99' engineering team, and how agentic workloads are changing the future of retrieval.

Dylan Patel Explains the AI War While Cooking | In-Context Cooking

Dylan Patel Explains the AI War While Cooking | In-Context Cooking

Dylan Patel of SemiAnalysis discusses the AI arms race, highlighting the massive $200B+ hyperscaler capex, the true semiconductor bottlenecks shifting from data centers back to fabs, the geopolitical chess game surrounding Taiwan and TSMC, and Nvidia's strategic battle against vertical integration.