Lenny's Podcast

Anthropic’s first technical PM on token maxing, the jagged edge, and living in the future | Dianne Penn

with Diane Penn
26 Jul 2026 5 min read 1h 10m

Diane Penn, Anthropic's first technical PM, argues that Opus 3's coding focus in early 2024 was the hidden inflection point that differentiated Claude — and that Opus 45 only succeeded because Claude Code existed as a 'vehicle' for the model's intelligence. She frames evals as 'the new PRDs' and says the best PMs right now should be asking: if Claude 8 comes out, what changes in what users do, and is what I'm building today forward-compatible with that?

Diane Penn
“Evals are the new PRDs.”
Diane is describing how the product role is changing at Anthropic, where evaluation frameworks have replaced traditional product requirements documents as the core driver of user value.
▶ 1:00
Diane Penn
“In 2023 when I started um nobody said anthropic and claude and coding in the same sentence.”
Diane is describing how coding was not associated with Claude or Anthropic in the early days, before they made it a deliberate training focus.
▶ 11:11
Diane Penn
“Opus 45 wouldn't have had that moment without a product like Cloud Code and Cloud Code I think wouldn't have had that type of adoption accelerated without Opus45.”
Diane is explaining why Opus 45 was a major inflection — not just the model itself, but the combination of model and product experience.
▶ 13:38
Diane Penn
“Unless you have the eval unless you have the systems to test um these jumps might actually happen and you don't know.”
Diane is explaining how emerging capabilities in AI models can appear discontinuously — and why evals are essential to catch capabilities you didn't know the model had developed.
▶ 19:09
Diane Penn
“Let's say claude 8 comes around what do what changes in what users do and then what should what does that mean for how you're building today?”
Diane shares the question she frequently asks her product team to force forward-compatible thinking given how rapidly models are improving.
▶ 34:44
Diane Penn is Head of Product for the AI Research and Labs teams at Anthropic, where she joined as the company's first technical product manager over three years ago when the product team had just five engineers. She has helped ship every model Anthropic has released, from Claude 2 through Fable, and has incubated and launched major products including Claude Code, MCP, Skills, Claude Design, and core capabilities like computer use, tool use, and reasoning. Her unique position at the intersection of research and product gives her a front-row seat to how frontier AI models are trained, evaluated, and brought to users.
1
Evals are the new PRDs for AI PMs At Anthropic, the team has replaced traditional product requirements documents with evaluation frameworks as the primary way to drive user value and guide model improvement. If you're building AI products, your ability to write precise, actionable evals is now more important than writing feature specs.
2
Ask 'what does Claude 8 change?' before shipping Diane's standing question to her team is: if Claude 8 arrives, what changes in what users do — and is what we're building today forward-compatible with that future? This forces product decisions to be durable across model generations rather than optimized for today's capabilities.
3
Model plus product vehicle — both required Opus 45 only became a breakout moment because Claude Code existed as a 'vehicle' to deliver its intelligence to users. Diane's framing: 'you need frontier products in order to have frontier models' resonate. A powerful model without the right product experience won't achieve adoption, and vice versa.