Emergent behavior

Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs

Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs

Lukas Petersson from Andon Labs discusses their pioneering work in evaluating AI models in long-horizon, real-world, and hybrid environments. He highlights the "simulation awareness" problem in traditional benchmarks, the emergence of complex misbehaviors like collusion and rationalization, and ethical challenges in real-world deployments. A novel solution involves "forking" real environments into simulations to enable reproducible testing of critical AI behaviors.

Securing the AI Frontier: Irregular Founder Dan Lahav

Securing the AI Frontier: Irregular Founder Dan Lahav

Dan Lahav, co-founder of Irregular, discusses the future of "frontier AI security," a proactive approach for a world where AI models are autonomous agents. He explains how emergent behaviors, such as models socially engineering each other or outmaneuvering traditional defenses like Windows Defender, signal a major paradigm shift. Lahav argues that as economic activity shifts to AI-on-AI interactions, traditional security methods like anomaly detection will break down, forcing enterprises and governments to rethink defense from first principles.