Skip to content
Artwork for Raising An Agent
Raising An Agent · Aug 21, 2025 · 53 min

Episode 8

In this episode of Raising an Agent, Beyang and Camden dive into how the Amp team evaluates models for agentic coding. They break down why tool calling is the key differentiator, what went wrong with Gemini Pro, and why open models like K2 and Qwen are promising but not ready as main drivers. They share first impressions of GPT-5, explore the idea of alloying models, and explain why qualitative "vibe checks" often matter more than benchmarks. If you want to understand how Amp thinks about model selection, subagents, and the future of coding with agents, this episode has you covered.

0:00-53:15

transcript

No transcript — this publisher did not publish one.

show notes

In this episode of Raising an Agent, Beyang and Camden dive into how the Amp team evaluates models for agentic coding. They break down why tool calling is the key differentiator, what went wrong with Gemini Pro, and why open models like K2 and Qwen are promising but not ready as main drivers. They share first impressions of GPT-5, explore the idea of alloying models, and explain why qualitative "vibe checks" often matter more than benchmarks. If you want to understand how Amp thinks about model selection, subagents, and the future of coding with agents, this episode has you covered.

more episodes

All episodes