Claude

Continual System Prompt Learning for Code Agents – Aparna Dhinakaran, Arize

Continual System Prompt Learning for Code Agents – Aparna Dhinakaran, Arize

The talk by Aparna Dhinakaran introduces "system prompt learning" as an efficient alternative to traditional Reinforcement Learning for improving large language model-based coding agents. By leveraging LLM-as-a-judge evaluations to generate English feedback and explanations for code failures, agents can automatically refine their system prompts and rules. This method, demonstrated on Claude and Klein, significantly boosts performance on benchmarks like SWEBench with minimal data, highlighting the critical role of high-quality evaluation prompts.

Disney's AI bet: USD 1B OpenAI content deal explained

Disney's AI bet: USD 1B OpenAI content deal explained

Experts Tim Hwang, Marina Danilevsky, Martin Keen, and Kush Varshney discuss Disney's partnership with OpenAI, Time Magazine's 'Architects of AI' Person of the Year, NVIDIA's Nemotron 3 model release, and the implications of Anthropic's leaked 'Soul Document' for model alignment and the future of prompting.

Anthropic stops AI spies, the new OWASP Top 10 and the rise of small-time ransomware

Anthropic stops AI spies, the new OWASP Top 10 and the rise of small-time ransomware

Experts discuss a report from Anthropic on a nearly autonomous AI-driven espionage campaign, debating its significance. The conversation explores the rise of agentic AI in attacks, the new 2025 OWASP Top 10, the fragmentation of the ransomware landscape, and the role of cyber insurance as a de facto regulator.

How Claude is transforming financial services

How Claude is transforming financial services

Anthropic's team discusses Claude for Financial Services, an agentic AI solution designed to transform financial workflows. They explore how Claude's core strengths in coding and reasoning are applied to tasks like real-time data analysis and generating investor-ready reports, highlighting practical customer examples and future developments.

Introducing Claude for Life Sciences

Introducing Claude for Life Sciences

Anthropic's Jonah Cool and Eric Kauderer-Abrams outline their vision for making Claude an indispensable AI research assistant for scientists. They discuss a multi-faceted strategy that includes enhancing model capabilities for long-horizon tasks, building a rich ecosystem through partnerships with companies like Benchling and 10x Genomics, and applying Claude across the entire R&D lifecycle—from bioinformatics analysis to navigating regulatory submissions.

Building with MCP and the Claude API

Building with MCP and the Claude API

A discussion with Anthropic engineers Alex Albert, John Welsh, and Michael Cohen about the Model Context Protocol (MCP). They cover its origins as an open standard, best practices for tool design and prompt engineering, and the future of the ecosystem where high-quality MCP servers will become a key competitive advantage.