Ai agents

CAN CHINA BEAT WAYMO?

CAN CHINA BEAT WAYMO?

This episode discusses three critical topics in AI: the true nature of recent AI agent "breakouts," arguing they highlight governance and security flaws rather than model danger; the role of AGI narratives in fueling the current AI investment bubble and questioning its sustainability; and China's aggressive strategy in the global robotaxi market, potentially outpacing Western counterparts like Waymo.

The New Primitives: Building AI Native Software — Kwindla Kramer, Daily

The New Primitives: Building AI Native Software — Kwindla Kramer, Daily

Kwindla Hultman Kramer argues that current AI agents are akin to 1995 web pages – a foundational primitive, not the ultimate destination. Drawing a historical parallel, he predicts the emergence of "AI native software" that will build upon and surpass agents, much like web applications evolved from simple web pages. He illustrates this through computing history, highlighting transformative shifts like VisiCalc's impact on accounting and the vision of Apple's Knowledge Navigator, and concludes by showcasing a game, Gradient Bang, that demonstrates the core primitives of this future AI native software.

Garry Tan: "Personal AGI Is How You Stay Under Your Own Power"

Garry Tan: "Personal AGI Is How You Stay Under Your Own Power"

YC President Garry Tan introduces "Personal AGI," a concept where individuals own and train AI agents on their unique context, transforming personal productivity and startup creation. He shares his workflow, advocating for Markdown as code and emphasizing the critical importance of owning one's intellectual assets in the age of AI. The talk outlines a practical 5-step guide to building a personal AI system and explores the profound implications for innovation and individual empowerment.

The State of Model Routing — NVIDIA, Cognition, OpenRouter

The State of Model Routing — NVIDIA, Cognition, OpenRouter

A deep dive into model routing strategies for AI/ML production, featuring experts from Cognition, OpenRouter, and NVIDIA. Key topics include optimizing costs with multi-model systems, delegating tasks between frontier and smaller models, managing context efficiently (sidekicks, compaction), and adapting to dynamic task complexities. The panel discusses the fragility of naive routing, the cost implications of in-distribution vs. out-of-distribution tasks, and the evolution of auto-routers driven by real-world usage patterns like OpenClaw's heartbeats. Insights also cover NVIDIA's Flex Run for dynamic model sizing, hallucination probes for detecting model limitations, and the future of hybrid local/cloud routing and model collaboration.

Agentic Engineering vs Software Engineering: Beyond Vibe Coding

Agentic Engineering vs Software Engineering: Beyond Vibe Coding

Anna Gutowska explains the paradigm shift in software engineering towards "agentic engineering," where AI agents execute goals defined by developers. She differentiates this from traditional, AI-assisted, and vibe coding, highlighting the increased importance of human oversight, orchestration, and verification in a world of probabilistic AI systems, and how this redefines the developer's role.

Rethinking Environments for Long-Horizon Work — Rayan Garg, Theta Software

Rethinking Environments for Long-Horizon Work — Rayan Garg, Theta Software

Rayan Garg from Theta Software delves into the complexities of defining and evaluating "long horizon" tasks for AI agents. He critiques current metrics and benchmarks, emphasizing the critical role of sophisticated environment design and robust verifiers (judge models) in driving true progress, particularly in "software-failing domains." The discussion highlights issues like task ambiguity, state changes, and the necessity for granular reward signals for effective model training.