Ai safety

Inside Mythos: Anthropic's Locked-Down Frontier Model — with Jon Krohn (@JonKrohnLearns)

Inside Mythos: Anthropic's Locked-Down Frontier Model — with Jon Krohn (@JonKrohnLearns)

Anthropic's Claude Mythos Preview is a frontier AI model with emergent hacking capabilities so advanced it's being withheld from public release. This summary details its near 100x performance leap in exploit generation, the 'Project Glasswing' industry consortium for responsible disclosure, and practical advice for developers to secure AI-generated code in this new era of automated vulnerability discovery.

Waymo's Dmitri Dolgov: 20 Million Rides and the Road to Full Autonomy

Waymo's Dmitri Dolgov: 20 Million Rides and the Road to Full Autonomy

Dmitri Dolgov, co-CEO of Waymo, discusses the 20-year journey from the DARPA challenge to full autonomy. He explains the Waymo Foundation Model—a multimodal world action model powering the driver, simulator, and critic—and how their "end-to-end plus" architecture enables superhuman safety and exponential scaling.

What Do Models Still Suck At? - Peter Gostev, Arena.ai, BullshitBench

What Do Models Still Suck At? - Peter Gostev, Arena.ai, BullshitBench

Despite benchmarks showing relentless progress, many users remain dissatisfied with LLM responses in real-world scenarios. This summary explores two key analyses—a custom 'nonsense question' benchmark and trends from Chatbot Arena's 'dislike both' data—to reveal the persistent gaps in model reasoning, reliability, and domain-specific understanding.

Claude Opus 4.7, Apple’s AI glasses and Allbirds AI pivot

Claude Opus 4.7, Apple’s AI glasses and Allbirds AI pivot

Experts analyze Anthropic's surprise release of Claude 4.7, speculating it's a distilled version of the Mythos model. The discussion also covers Apple's new three-pronged AI wearables strategy, a Gallup poll showing rising but incremental AI adoption in the workplace, and DeepMind's research into harmful AI manipulation.

Cognitive Exhaust Fumes, or: Read-Only AI Is Underrated — Šimon Podhajský, Head of AI, Waypoint

Cognitive Exhaust Fumes, or: Read-Only AI Is Underrated — Šimon Podhajský, Head of AI, Waypoint

A deep dive into a "read-only" personal AI system that analyzes your digital footprint—or "cognitive exhaust fumes"—from sources like email, notes, and browsing history. The author argues that this observer approach provides more profound insights and is inherently safer than action-oriented AI agents, by preventing data contamination and mitigating the high-stakes risks of write-access errors.

Episode 15 - Inside the Model Spec

Episode 15 - Inside the Model Spec

OpenAI researcher Jason Wolfe explains the Model Spec, the public framework defining intended model behavior. This summary covers its core principles like the 'chain of command,' how it handles complex edge cases, its evolution through public feedback, and its future role in an increasingly autonomous AI landscape.