Cognition President of New Enterprise Jeff Wang envisions a future where AI agents prove their own work
In this episode
As AI agents run longer and produce more code, manual review will become the bottleneck. Cognition President of New Enterprise Jeff Wang is building a future where that’s no longer the case. In this episode, Jeff joins 1Password CTO Nancy Wang and special guest co-host Richard Liu, Head of API Products at Anthropic, for a conversation about cloud-based agents and how their ability to prove their own work could shift human review from checking every line of code to auditing outcomes, enabling greater agent autonomy while preserving accountability and trust.
Note: this episode was filmed in April of 2026 and does not reflect the advancements that have occurred since then.
How agents prove they finished the job
Agents need a method to prove that their work is correct
Event-driven agents are absorbing the grunt work nobody wants
Playbooks are the most overlooked and underinvested piece of knowledge infrastructure
Agents should run in isolated sandboxes with only the specific access they need
Code reviews will eventually be replaced by proof that the code does what it’s supposed to
Also available in audio format on the following platforms: Apple Podcasts, Spotify
Jeff Wang
Jeff Wang is the President of New Enterprise at Cognition and former CEO of Windsurf (now part of Cognition), the company behind the first agentic IDE. Windsurf has scaled to serve a large and active community of developers, industry partners, and enterprise customers. Prior to Windsurf, he co-founded RocketFuel Education and RNR Capital, held executive roles at several startups, and began his career at Salesforce and Cisco.


Richard Liu
Richard Liu is Head of API Products at Anthropic, where he works on the API layer that powers Claude for enterprise customers, focusing on security, enterprise readiness, and the identity infrastructure behind products like the Claude Agent SDK and Claude Managed Agents. Earlier in his career, he held roles at Google, Meta, and AWS.


Get the episode transcript
More episodes

OpenAI Agent Security Lead Fotis Chantzis discusses one of the biggest unsolved problems in AI

Vercel Chief Product Officer Tom Occhino reflects on React and how AI transforms software creation

Cursor Head of Security Travis McPeak considers how to secure agents without slowing work

Braintrust founder and CEO Ankur Goyal examines why AI learning happens after launch

Mercor co-founder and co-CEO Adarsh Hiremath addresses the AI onboarding problem
