GenAI PM
person49 mentions· Updated Sep 8, 2026

Claire Vo

AI/PM creator credited with recapping the How I AI episode about Stripe Kai. Mentioned as the source of the newsletter item.

Key Highlights

  • Claire Vo is repeatedly cited for practical demonstrations of AI agents handling real operational and product workflows.
  • She ran a live benchmark across seven models on PM-relevant tasks like PRDs, prototyping, bug triage, and agentic coding.
  • Her posts surface important AI PM concerns including prompt injection, connector transparency, multi-account UX, and reliability.
  • She shared concrete Codex use cases spanning accounting, inbox triage, QA, security questionnaires, and SaaS setup.
  • She was credited with recapping the How I AI episode on Stripe Kai, Stripe’s internal company-brain platform.

Claire Vo

Overview

Claire Vo is an AI product and workflow creator frequently cited in discussions about practical, operator-level uses of frontier AI tools. In the newsletter corpus, she appears as both a source and a demonstrator: recapping How I AI episodes, publishing hands-on experiments with agents and coding tools, and sharing concrete examples of how AI systems can be used for operational work, benchmarking, and internal automation.

For AI Product Managers, Claire Vo matters because her mentions consistently sit at the intersection of product strategy, agent UX, and real-world execution. Rather than treating AI as abstract capability, her posts and appearances focus on implementation details: model comparison, tool ergonomics, internal agent design, operational delegation, connector trust, and safety concerns like prompt injection and data handling. That makes her a useful signal for PMs tracking how advanced users actually evaluate and deploy AI-native workflows.

Key Developments

  • 2026-07-24: Claire Vo demonstrated handing more computer-based work to GPT-5.6-powered Codex agents, including front-end bug detection, synthetic browser persona research, LinkedIn inbox cleanup, and personal shopping.
  • 2026-07-25: She ran a live How I AI benchmark comparing seven AI models across six tasks such as PRD creation, prototyping, wireframing, bug triage, agentic coding, and agent voice; Opus 5 finished first.
  • 2026-08-04: Claire Vo recommended @evedev_ as a default framework for internal agents, citing instructions, skills, channels, and connectors; she also pointed to a forthcoming How I AI episode on building a PR review and approval agent.
  • 2026-08-08: She showed how she used Codex to build an always-on toolbar app on her Mac for managing a smart lightbulb.
  • 2026-08-12: Claire Vo praised @bot’s UX, especially multi-account sign-in for tools like Slack and Google Workspace, framing that capability as valuable for managing systems across multiple businesses.
  • 2026-08-16: She described her own multi-agent setup: two main OpenClaws, one “lifeguard” OpenClaw, and one Codex configured via SSH to recover the others, calling it difficult but worthwhile to maintain.
  • 2026-08-20: Claire Vo shared operational use cases for Codex browser/Chrome/computer, including accounting, inbox management, Stripe Radar configuration, browser QA, security questionnaires, SaaS setup when APIs are unavailable, and subscription cancellation.
  • 2026-08-24: She raised questions about whether agents can work without indexing source data, how ephemeral connectors versus stored data should be disclosed to users, and how to mitigate prompt injection; she also criticized immature evaluation approaches that lacked user empathy.
  • 2026-09-06: Claire Vo shared that an agent she named Penny Pincher was negotiating a vintage Rolex roughly 10 messages into the interaction.
  • 2026-09-08: She was credited with recapping a How I AI episode featuring Sharadh Krishnamurthy on Stripe Kai, described as Stripe’s “company brain,” including themes such as projects-as-governance, skill routing and telemetry, and a skills platform serving roughly 10,000 teammates.

Relevance to AI PMs

1. Practical agent design patterns: Claire Vo’s examples show how agents move from novelty to utility when applied to concrete workflows like QA, inbox management, approvals, and configuration tasks. PMs can use these examples to identify high-friction, browser-native work that is ripe for delegation.

2. Evaluation and benchmarking discipline: Her live model benchmark is a reminder that PMs need task-based evaluation, not just benchmark scores. Comparing models across PRDs, prototyping, bug triage, and coding helps teams map the right model to the right product job.

3. Trust, UX, and governance considerations: Her comments on prompt injection, connector behavior, multi-account sign-in, source-data indexing, and rescue setups highlight the product requirements that matter after the demo: safety, transparency, account context, recoverability, and operational reliability.

Related

  • How I AI / How I AI Podcast: A major context for Claire Vo’s presence; she appears as a benchmark host, recap source, and commentator on AI workflows.
  • ChatPRD: Mentioned in one alias string and likely central to her builder identity, connecting her to AI-assisted product management workflows.
  • Codex / GPT-5.6 / OpenAI: Frequently associated with her demonstrations of browser agents, computer-use automation, and lightweight app creation.
  • OpenClaw: Connected through her multi-agent operational setup, including backup and rescue architecture.
  • EveDev / internal agents: Tied to her recommendations for internal agent frameworks with built-in skills, instructions, and connectors.
  • Stripe / Sharadh Krishnamurthy / Kai: Connected through her recap of the How I AI episode on Stripe’s internal “company brain.”
  • Anthropic / Claude / Opus models / Gemini: Relevant through her model comparisons and the broader ecosystem of tools she evaluates.
  • Slack, Google Workspace, SSH, Stripe Radar: Examples of the real business systems and interfaces around which her AI workflow commentary is grounded.

Newsletter Mentions (49)

2026-09-08
How 1.5 engineers got Stripe Kai in two weeks #1 in Claire Vo recapped a How I AI episode in which Sharadh Krishnamurthy demonstrated Kai, Stripe’s “company brain,” which the post says 1.5 engineers and 2 weeks helped deliver.

#1 in Claire Vo recapped a How I AI episode in which Sharadh Krishnamurthy demonstrated Kai, Stripe’s “company brain,” which the post says 1.5 engineers and 2 weeks helped deliver. The episode covered projects as governance, skill routing and telemetry, and a skills platform described as working for 10k teammates.

2026-09-06
claire vo shared that she has an agent named Penny Pincher, which, about 10 messages into their interaction, was negotiating a vintage Rolex for her.

#8 𝕏 claire vo shared that she has an agent named Penny Pincher, which, about 10 messages into their interaction, was negotiating a vintage Rolex for her.

2026-08-24
#6 𝕏 claire vo 🖤 questioned whether agents can work without indexing source data, how ephemeral connectors and stored data should be presented to users, and how to prevent prompt injection.

#6 𝕏 claire vo 🖤 questioned whether agents can work without indexing source data, how ephemeral connectors and stored data should be presented to users, and how to prevent prompt injection. She characterized an unnamed harness as immature and lacking user-empathetic evaluation, while noting that self-serve data deletion shipped but promised email follow-up did not occur.

2026-08-20
Claire Vo shared how she uses Codex browser/Chrome/computer for operational tasks including accounting, inbox management, Stripe Radar configuration, browser-based QA, security questionnaires, SaaS setup when an API is unavailable, and subscription cancellation.

GenAI PM Daily August 20, 2026 GenAI PM Daily 🎧 Listen to this brief 3 min listen Today's top 20 insights for PM Builders, ranked by relevance from Blogs, X, YouTube, and LinkedIn. OpenAI announces Zero Data Retention for frontier models #1 📝 OpenAI News Offering Zero Data Retention for frontier models - OpenAI announces offering zero data retention for frontier models, committing to not retain user data for those models and clarifying how this impacts customers and data handling. The post outlines the company's privacy-focused approach for frontier model interactions. Also covered by: @OpenAI , @OpenAI , @Sam Altman #2 𝕏 Cursor announced that it can now monitor pull requests, watch a Slack thread, and run scheduled tasks. Cloud agents automatically subscribe to pull requests they create and drive them to completion. #3 𝕏 Mustafa Suleyman announced that MAI-Image-2.5 is ranked #1 on the Artificial Analysis leaderboard for image editing. #4 𝕏 Logan Kilpatrick announced that Google AI Studio now supports GitHub repository imports and bi-directional push/pull synchronization. A new UI also supports force pushes and merges. #5 𝕏 Qwen shared that Qwen3.8-27B ranked as the #1 open-weight model on Harvey’s Legal Agent benchmark, describing it as capable of professional tasks while remaining small enough to run locally. #6 𝕏 NVIDIA shared that NVIDIA cuOpt, its open-source solver, is the fastest open-source solver on Hans Mittelmann benchmarks across three optimization problem classes. #7 𝕏 Results from benchmarks of 300+ NVIDIA verified skills on real tasks showed that using skills improved correctness by 41 points, effectiveness by 39 points, and efficiency by 35 points. SkillEvaluator is open source for testing skills before shipping. #8 𝕏 Philipp Schmid shared that Gemini 3.7 Flash ranked first on Artificial Analysis’s new AA-AnalystAgent, which covers 80 real-world quantitative analysis tasks across 14 business and scientific domains. #9 𝕏 Claire Vo shared how she uses Codex browser/Chrome/computer for operational tasks including accounting, inbox management, Stripe Radar configuration, browser-based QA, security questionnaires, SaaS setup when an API is unavailable, and subscription cancellation.

2026-08-16
claire vo 🖤 commented that she maintains two main OpenClaws, one lifeguard OpenClaw, and one Codex tuned to connect via SSH and rescue both main OpenClaws—a setup she calls annoying to maintain and a labor of love.

#8 𝕏 claire vo 🖤 commented that she maintains two main OpenClaws, one lifeguard OpenClaw, and one Codex tuned to connect via SSH and rescue both main OpenClaws—a setup she calls annoying to maintain and a labor of love.

2026-08-12
"#12 𝕏 claire vo 🖤 praised @bot’s UX, highlighting multi-account sign-in for services such as Slack and Google Workspace as its key feature for managing systems across multiple businesses."

#12 𝕏 claire vo 🖤 praised @bot’s UX, highlighting multi-account sign-in for services such as Slack and Google Workspace as its key feature for managing systems across multiple businesses. She said she tested @bot early and provided feedback, though it has not yet replaced another tool she represented with a lobster emoji.

2026-08-08
claire vo demonstrated how she used Codex to turn something into an always-on toolbar app for managing her smart lightbulb from her Mac, in a post referencing a video.

#12 𝕏 claire vo demonstrated how she used Codex to turn something into an always-on toolbar app for managing her smart lightbulb from her Mac, in a post referencing a video.

2026-08-04
claire vo recommends @evedev_ as a default framework for internal agents, citing its instructions, skills, built-in channels, and connectors.

#11 𝕏 claire vo recommends @evedev_ as a default framework for internal agents, citing its instructions, skills, built-in channels, and connectors. She also announced a forthcoming How I AI episode about building a PR review and approval agent with Eve, while a quoted post describes Vercel’s internal AI agent @v as powered by @evedev_.

2026-07-25
#17 ▶️ I hate Opus 5. It’s the best model, anyway. How I AI Podcast Claire Vo runs a live How I AI benchmark comparing seven AI models (Opus 5, Sonnet 5, Fable, Opus 4, Mabu, GPT Terra and Gemini 3.1 Pro) across six tasks—PRD creation, prototype creation, wireframe creation, bug triage, agentic coding and agent voice—scored 70% by her manual vibe check and 30% by GPT-5.5, with Opus 5 emerging first on the leaderboard.

#17 ▶️ I hate Opus 5. It’s the best model, anyway. How I AI Podcast Claire Vo runs a live How I AI benchmark comparing seven AI models (Opus 5, Sonnet 5, Fable, Opus 4, Mabu, GPT Terra and Gemini 3.1 Pro) across six tasks—PRD creation, prototype creation, wireframe creation, bug triage, agentic coding and agent voice—scored 70% by her manual vibe check and 30% by GPT-5.5, with Opus 5 emerging first on the leaderboard. The benchmark evaluated seven models—Opus 5, Sonnet 5, Fable, Opus 4, Mabu, GPT Terra and Gemini 3.1 Pro—in a blind test over six tasks: PRD creation, prototype creation, wireframe creation, bug triage, agentic coding and agent voice.

2026-07-24
claire vo is retiring her keyboard by handing her computer to GPT-5.6–powered Codex agents, demoing automated front-end bug detection (03:49), synthetic browser persona research (24:26), LinkedIn inbox cleanup (41:19) and personal shopping (44:39).

#23 𝕏 claire vo is retiring her keyboard by handing her computer to GPT-5.6–powered Codex agents, demoing automated front-end bug detection (03:49), synthetic browser persona research (24:26), LinkedIn inbox cleanup (41:19) and personal shopping (44:39). #24 📝 Ampcode Chronicle Event Driven Orbs - Amp orbs can now be woken by external HTTP requests: amp.createWebhook registers a durable webhook endpoint that verifies signatures, deduplicates deliveries, and spawns read-only orb threads with trusted repository/event/actor metadata so the orb can inspect and act on GitHub issues, CI failures, Linear issues, Discord messages, etc.

Related

Anthropiccompany

An AI company whose Threat Intelligence team published a report on misuse of Claude and related countermeasures. The newsletter highlights evolving malicious-use patterns and defensive responses.

Claude Codetool

Anthropic’s coding agent. It is relevant to AI PMs as a coding workflow product competing in enterprise and community adoption.

OpenAIcompany

An AI company that released the Agents API and GPT-Live-1, both aimed at helping builders ship production-grade agent and voice experiences. It is also discussed in relation to GPT-6 Astra, benchmarking, and evidence tracing features.

Claudetool

Anthropic's AI assistant/model referenced in a threat-intelligence report about misuse attempts. The report discusses cases, disruptions, and countermeasures over eight months of activity.

Guillermo Rauchperson

Founder and CEO of Vercel, often sharing product, pricing, and infrastructure updates. Here he recaps Vercel price cuts and AI Gateway token growth.

Cursortool

An AI coding tool that introduced Projects, a persistent coordinator-agent workflow. The feature moves teams away from task-by-task chats toward a single long-running thread with subagents.

Peter Yangperson

Person mentioned sharing a tutorial and a set of product principles in the newsletter. He is presented as a creator/commentator in AI product content.

Codextool

OpenAI’s coding tool/agent used for software development workflows. It matters for PMs as a replacement or alternative in enterprise coding adoption.

Lenny Rachitskyperson

Newsletter and podcast personality who recapped a discussion on AI job impacts and competition in the AI stack. He is cited as the source of the summary in the newsletter.

Vercelcompany

A developer platform and hosting company with a growing AI product surface, including v0 and AI Gateway. The newsletter cites product updates, pricing changes, and usage growth across its AI infrastructure offerings.

ChatGPTtool

OpenAI’s conversational AI product and plugin ecosystem. In this newsletter it is the target platform for a LlamaParse connector in the plugin directory.

OpenClawtool

A Slack-connected setup or workspace mentioned as being configured using AsideAI. It is relevant as an example of rapid AI-assisted integration setup.

Greg Isenbergperson

An entrepreneur and creator featured in a segment about making money with a Grok bot workflow. He is associated here with commentary on AI-driven newsletter operations.

MCPconcept

A protocol for connecting agents to external tools and systems in a standardized way. The newsletter mentions setup instructions that can be pasted into an agent to configure MCP.

v0tool

Vercel’s AI app-building tool, used here for adding integrations and distributing changes through a changelog. It is relevant to PMs building AI-powered product experiences and web app workflows.

Slacktool

A workplace messaging and collaboration platform. In this newsletter it appears as an integration target for AI setup and automation.

Stripecompany

Financial infrastructure company mentioned as the builder and deployer of Kai. The newsletter highlights its internal AI system as an example of shipping useful AI tooling quickly.

GPT-5.5tool

A model used as an automated judge in Claire Vo’s benchmark. It contributes 30% of the scoring alongside her manual evaluation.

AI agentsconcept

Autonomous or semi-autonomous AI systems that can plan and take actions across tools and workflows. This is a core AI PM concept central to product design and evaluation.

Vercel AI Gatewaytool

A gateway for routing AI requests and coding agents with observability and budget controls. In this newsletter it is used as the target for a coding-agents setup command.

Figmacompany

A collaborative design platform referenced as an example of broad enterprise SaaS that may remain resilient in the AI era. It is contrasted with niche single-purpose products.

GPT-5.6tool

A frontier model release referenced as improving price-performance for developers. It is discussed as being available in Kiro for more cost-effective application development.

Claude Designtool

A Claude-based design workflow or surface that connects designs to v0. It matters for AI PMs as a design-to-app handoff layer.

Claude Opus 4.6tool

A Claude model version referenced as part of a prompt-comparison analysis. It serves as one endpoint for examining changes in Anthropic’s system prompt evolution.

prompt injectionconcept

A security risk where malicious instructions manipulate agent behavior. The newsletter references it in the context of skill/repo scanning tools.

chatprdtool

An AI-first product management tool or startup referenced by Claire Vo. The newsletter uses it in a discussion of shipping an AI-first version of an app without traditional PM tooling.

GPT 5.4tool

A GPT model variant used here for scientific reasoning and agentic chemistry experimentation. The newsletter frames it as a model capable of proposing experimental improvements and driving benchmarked workflows.

Opus 4.7tool

A Claude model variant referenced in Anthropic's cybersecurity evaluation report. It is one of the models involved in the incidents described.

Zapiercompany

Zapier provides automation workflows and connectors used to link Claude with Google Analytics in the tutorial. It appears here as an integration layer for LLM-powered business analytics.

Google Workspacecompany

Google's suite of productivity applications used for email, documents, spreadsheets, and calendaring. It is mentioned here as the environment Cursor agents can now operate across.

Eve.devtool

A framework for internal agents that emphasizes instructions, skills, channels, and connectors. It is presented as a default choice for building internal agent systems.

Midjourneytool

A generative media company referenced as an example of a public Discord-based workflow. It is used here to support the idea that visible communities can accelerate learning and product adoption.

Granolacompany

An AI meeting-transcript tool used as a source of meeting notes and context in the Claude Cowork workflow.

Gemini 3 Protool

A Gemini model variant used in a real workflow library project. The newsletter mentions it as one of the tools used to build the ChatPRD index.

Gemini 3.1 Protool

Google's latest Gemini model highlighted for improved reasoning and multimodal capabilities. It is positioned as a model that can code full environments and work with integrated generative audio and UI controls.

How I AItool

A media and podcast brand covering practical AI workflows and agent use cases. It appears here as the source of an upcoming episode and a cited podcast discussion.

Wade Fosterperson

CEO of Zapier who shares his personal AI stack and recruiting workflows. He is highlighted again in a YouTube segment about using AI inside company culture.

YouTubecompany

The video platform mentioned for its new Inspiration feature, which is criticized here as AI-generated slop.

Intercomcompany

A customer service software company that used Claude Code to improve engineering throughput. Relevant here for measuring AI adoption, productivity, and workflow instrumentation.

Agencytool

A PM capability emphasizing initiative and the ability to drive outcomes independently. In AI product management, it suggests using AI to amplify decision-making and execution.

Jenny Wenperson

Head of design at Claude, cited in the newsletter for discussing how AI tools are changing the design process. She is associated with Anthropic's design workflow.

GPT Imagestool

OpenAI's image generation tool referenced in a workflow for building landing pages, slides, and brand kits. It is used alongside Claude Design for content and brand asset creation.

Stay updated on Claire Vo

Get curated AI PM insights delivered daily — covering this and 1,000+ other sources.

Subscribe Free