Measuring the frontier
The OpenAI Podcast · 16 Jun 2026 · Episode 21
About this episode
The old tests are getting too easy. Tejal Patwardhan, who leads OpenAI's frontier evals team, talks with host Andrew Mayne about why evals matter for research, how benchmarks break or get gamed, and what models should be judged on next as they keep getting more capable.
About the show
The OpenAI Podcast — OpenAI. Long-form conversations with the people building at OpenAI — research, ChatGPT, Codex, infrastructure and where the industry is heading.
Official episode page
Watch on YouTube
More from The OpenAI Podcast
What racing reveals about working with AI
How a reasoning model cracked an 80-year-old math problem
Inside image generation's Renaissance moment
All listening
Browse the site
Home
about
story
work
expertise
ai
ai ai product development
ai ai agents
ai ai automation
ai ai consulting
ai arabic ai products
ai kuwait
toolkit web
toolkit claude
toolkit lovable
toolkit notion
toolkit webflow
toolkit shopify
toolkit wordpress
toolkit ai solutions
services
services business strategy
services growth planning
tools
blog
listening
stack
contact
quote
privacy
terms