GenAI PM
person40 mentions· Updated Sep 10, 2026

Santiago

An AI engineering voice in the newsletter, sharing practical systems and workflows for agents and specialized intelligence. He is cited on model selection, evals, serving, and agent setup.

Key Highlights

  • Santiago is a recurring AI engineering voice focused on specialized intelligence, agent workflows, model routing, and production system design.
  • He introduced the “value per token dollar” metric as a practical way to measure whether agent usage is economically viable.
  • His commentary emphasizes shifting from line-by-line review of AI-generated code toward stronger system-level verification.
  • He is associated with launches including Hyperagent cloud agents, BAND for agent communication, and an Apify Store skill for autonomous paid tool use.
  • For AI PMs, his work is most useful as a playbook for moving from AI demos to reliable, economically grounded agent products.

Santiago

Overview

Santiago is a recurring AI engineering voice in the GenAI PM ecosystem, often cited for practical thinking on how agent systems should actually be built, evaluated, routed, and operated. Across newsletter mentions, he shows up less as a pure commentator and more as a systems-minded builder focused on specialized intelligence: model selection, eval design, serving tradeoffs, agent setup, and workflow architecture.

For AI Product Managers, Santiago matters because his posts consistently translate frontier AI capability into product and operational decisions. His themes include measuring agent ROI, reducing manual code review in favor of system-level verification, designing agents that can use tools and payments autonomously, and building interfaces where agents act across email, browser, APIs, and cloud environments. That makes him a useful reference point for PMs working on agent products, AI-enabled developer tools, and applied model infrastructure.

Key Developments

  • 2026-06-30: Santiago proposed the “value per token dollar” ratio—value created divided by token cost—as a simple way to measure agent ROI. He noted that scores below 1 lose money, 1 breaks even, and above 1 is profitable, and credited matrix_build as an early adopter.
  • 2026-07-07: He proposed giving AI agents their own email addresses that users can CC, enabling agents to autonomously handle tasks while keeping human inboxes separate.
  • 2026-07-12: Santiago predicted AI video would evolve from static clips to real-time, interactive livestream-like experiences, pointing to early demos of personalized, dynamic media.
  • 2026-07-24: He launched a pay-as-you-go Apify Store skill that lets agents discover tools, handle HTTP 402 payment prompts, authorize USDC on Base, and execute Actors end-to-end inside agent workflows.
  • 2026-07-28: Under Santiago Hyperagent, he launched cloud-based agents with zero setup or servers required, featuring API skill learning, browser actions, code execution, image and video generation, and integrations with hundreds of tools.
  • 2026-07-30: Santiago introduced Kimi K3, describing it as an open-weight 2.8T-parameter coding model with 1M-token context, multimodal inputs, and long-running agentic sessions. He also highlighted the Verdent + Moonshot AI partnership behind its optimized coding workflows.
  • 2026-07-30: He also launched BAND, an interaction layer designed for personal agents to communicate across identities, channels, and routing, demonstrated through a calendar collaboration use case.
  • 2026-08-09: Santiago said he had stopped reading AI-generated code line by line and instead focused on verifying the overall system, arguing that expert human guidance still matters but exhaustive inspection does not scale.
  • 2026-08-20: He shared Open Bot, describing it as a free, open-source bot that works with any harness or model and can run anywhere.
  • 2026-08-25: Santiago recapped WAN 3.0, highlighting native 30-second video generation from text, images, or Omni Video, support for up to 20 reference assets including documents and webpages, and more consistent outputs with precise edits.
  • 2026-09-10: He shared four key engineering skills for specialized intelligence: choosing and shaping models, designing usage-reflective evals, serving under latency/cost/quality constraints, and routing requests while defining system policies.

Relevance to AI PMs

1. He provides practical frameworks for shipping agent products. Santiago’s posts emphasize model choice, eval quality, serving tradeoffs, and routing policy—core decisions PMs must make when moving from demo agents to production systems.

2. He offers useful operating metrics and decision models. The “value per token dollar” concept gives PMs a straightforward way to connect model usage to business outcomes, especially for agent workflows with variable tool use and long task chains.

3. He points toward emerging product patterns for agents. Examples like agent-specific email addresses, cloud agents with tool learning, cross-agent communication layers, and autonomous paid tool execution help PMs think beyond chatbots toward full workflow automation.

Related

  • hyperagent: Closely tied to Santiago through the launch of cloud-based agents and the “Santiago Hyperagent” alias.
  • ai-agents / llm-agents / coding-agents: Central categories connected to his commentary on setup, orchestration, verification, and long-running agent workflows.
  • merge-gateway / internal-ai-model-gateways: Relevant to his emphasis on routing requests, model selection, and policy definition across systems.
  • prompt-versioning / evals / context / continuous-learning: Connected to his focus on usage-reflective evaluation and improving specialized intelligence over time.
  • Open Bot / BAND / Apify Store / browser-automation-system / computer-use-agent: Examples of the agent capabilities and interaction patterns he promotes.
  • Kimi K3 / Moonshot AI / Verdent / WAN 3.0 / Fireworks AI: Technologies and companies he has highlighted while discussing coding models, video generation, and specialized model systems.
  • USDC / Base: Relevant through his autonomous payment workflow for agents buying and executing tools.
  • ai-generated-code / IDE / Claude Code / Cursor / Codex / Gemini CLI: Related to his viewpoint that the workflow focus is shifting from manual code inspection toward system verification and agent-native tooling.

Newsletter Mentions (39)

2026-09-10
Santiago shared four key engineering skills for “specialized intelligence”: choosing and shaping models, designing usage-reflective evals, serving models under latency, cost, and quality constraints, and routing requests and defining system policies.

#10 𝕏 Santiago shared four key engineering skills for “specialized intelligence”: choosing and shaping models, designing usage-reflective evals, serving models under latency, cost, and quality constraints, and routing requests and defining system policies. Fireworks AI’s flagship specialized intelligence conference, Forge ’26, is scheduled for November 3 at Pier 27 in San Francisco, and the post included an application link.

2026-08-25
Santiago recapped WAN 3.0’s ability to generate native 30-second videos from text, images, or Omni Video, use up to 20 reference assets including documents and webpages, and deliver more consistent outputs and precise edits with less regeneration.

GenAI PM Daily August 25, 2026 GenAI PM Daily 🎧 Listen to this brief 3 min listen Today's top 19 insights for PM Builders, ranked by relevance from Blogs, YouTube, and LinkedIn. GPT-5.6 in Kiro advances developer price-performance #1 📝 OpenAI News Advancing price-performance for developers with GPT‑5.6 in Kiro - Announces availability of GPT‑5.6 in Kiro to improve price-performance for developers, enabling more cost-effective and performant model access for applications. #4 𝕏 Santiago recapped WAN 3.0’s ability to generate native 30-second videos from text, images, or Omni Video, use up to 20 reference assets including documents and webpages, and deliver more consistent outputs and precise edits with less regeneration.

2026-08-20
Santiago shared that Open Bot is a free, open-source bot that works with any harness or model and can run anywhere.

GenAI PM Daily August 20, 2026 GenAI PM Daily 🎧 Listen to this brief 3 min listen Today's top 20 insights for PM Builders, ranked by relevance from Blogs, X, YouTube, and LinkedIn. OpenAI announces Zero Data Retention for frontier models #1 📝 OpenAI News Offering Zero Data Retention for frontier models - OpenAI announces offering zero data retention for frontier models, committing to not retain user data for those models and clarifying how this impacts customers and data handling. The post outlines the company's privacy-focused approach for frontier model interactions. Also covered by: @OpenAI , @OpenAI , @Sam Altman #2 𝕏 Cursor announced that it can now monitor pull requests, watch a Slack thread, and run scheduled tasks. Cloud agents automatically subscribe to pull requests they create and drive them to completion. #3 𝕏 Mustafa Suleyman announced that MAI-Image-2.5 is ranked #1 on the Artificial Analysis leaderboard for image editing. #4 𝕏 Logan Kilpatrick announced that Google AI Studio now supports GitHub repository imports and bi-directional push/pull synchronization. A new UI also supports force pushes and merges. #5 𝕏 Qwen shared that Qwen3.8-27B ranked as the #1 open-weight model on Harvey’s Legal Agent benchmark, describing it as capable of professional tasks while remaining small enough to run locally. #6 𝕏 NVIDIA shared that NVIDIA cuOpt, its open-source solver, is the fastest open-source solver on Hans Mittelmann benchmarks across three optimization problem classes. #7 𝕏 Results from benchmarks of 300+ NVIDIA verified skills on real tasks showed that using skills improved correctness by 41 points, effectiveness by 39 points, and efficiency by 35 points. SkillEvaluator is open source for testing skills before shipping. #8 𝕏 Philipp Schmid shared that Gemini 3.7 Flash ranked first on Artificial Analysis’s new AA-AnalystAgent, which covers 80 real-world quantitative analysis tasks across 14 business and scientific domains. #9 𝕏 Claire Vo shared how she uses Codex browser/Chrome/computer for operational tasks including accounting, inbox management, Stripe Radar configuration, browser-based QA, security questionnaires, SaaS setup when an API is unavailable, and subscription cancellation. #10 𝕏 Madhu Guru shared an eval strategy for AI products: define a rubric, use the best available measurement process to establish a trustworthy quality frontier, then reduce costs through automation, smaller judge models, sampling, and deterministic checks where relevant.

2026-08-09
#8 𝕏 Santiago said he had stopped reading AI-generated code and had not looked at any of it for two weeks, concluding that his time was better spent designing ways to verify the overall system.

GenAI PM Daily August 09, 2026 GenAI PM Daily 🎧 Listen to this brief 3 min listen Today's top 10 insights for PM Builders. Claude Code sessions can now message each other #8 𝕏 Santiago said he had stopped reading AI-generated code and had not looked at any of it for two weeks, concluding that his time was better spent designing ways to verify the overall system. He said experienced human guidance is still needed, but checking every line is not, and argued that the IDE is on its way to the graveyard and new tools are needed.

2026-07-30
Santiago introduced Kimi K3, an open‐weight 2.8T‐parameter coding model with 1 M‐token context, multimodal inputs and agentic long‐run sessions.

#5 𝕏 Santiago introduced Kimi K3, an open‐weight 2.8T‐parameter coding model with 1 M‐token context, multimodal inputs and agentic long‐run sessions. He highlighted the Verdent + Moonshot AI partnership behind its optimized agentic coding workflows. #8 𝕏 Santiago launched BAND, an interaction layer enabling personal agents to communicate across identities, channels, and routing, showcased with a calendar collaboration demo.

2026-07-28
Santiago Hyperagent launches cloud-based agents—zero setup or servers required—that learn new API skills, support browser actions, code execution, image/video generation and integrate with hundreds of tools.

GenAI PM Daily July 28, 2026. Santiago is associated with a product launch for cloud-based agents.

2026-07-24
Santiago launched a pay-as-you-go Apify Store skill that lets agents autonomously discover, trigger HTTP 402 payment prompts, authorize USDC on Base, and execute Actors end-to-end—bringing a full AI tools marketplace into agentic workflows.

#11 𝕏 Santiago launched a pay-as-you-go Apify Store skill that lets agents autonomously discover, trigger HTTP 402 payment prompts, authorize USDC on Base, and execute Actors end-to-end—bringing a full AI tools marketplace into agentic workflows. #12 𝕏 LlamaIndex 🦙 launched new Go and Java SDKs plus a CLI for parsing documents in just a few lines of code or straight from your terminal—grab an API key and start for free today.

2026-07-12
Santiago predicts AI video will shift from static clips to real-time, interactive livestream-style experiences (think Minority Report–style personalized ads) and shares a demo link showcasing this early potential.

#13 𝕏 Santiago predicts AI video will shift from static clips to real-time, interactive livestream-style experiences (think Minority Report–style personalized ads) and shares a demo link showcasing this early potential.

2026-07-07
Santiago proposes giving AI agents their own email addresses you CC on messages, so they can autonomously handle tasks while keeping your inbox separate.

GenAI PM Daily July 07, 2026 GenAI PM Daily 🎧 Listen to this brief 3 min listen Today's top 20 insights for PM Builders, ranked by relevance from Blogs, X, YouTube, and LinkedIn. #19 𝕏 Santiago proposes giving AI agents their own email addresses you CC on messages, so they can autonomously handle tasks while keeping your inbox separate.

2026-06-30
#12 𝕏 Santiago proposes the “value per token dollar” ratio—value created ÷ token cost—to measure agent ROI, where below 1 loses money, at 1 breaks even, and above 1 is profitable.

#12 𝕏 Santiago proposes the “value per token dollar” ratio—value created ÷ token cost—to measure agent ROI, where below 1 loses money, at 1 breaks even, and above 1 is profitable. He credits @matrix_build as the first team to adopt this metric.

Related

Claude Codetool

Anthropic’s coding agent. It is relevant to AI PMs as a coding workflow product competing in enterprise and community adoption.

Claudetool

Anthropic's AI assistant/model referenced in a threat-intelligence report about misuse attempts. The report discusses cases, disruptions, and countermeasures over eight months of activity.

Cursortool

An AI coding tool that introduced Projects, a persistent coordinator-agent workflow. The feature moves teams away from task-by-task chats toward a single long-running thread with subagents.

Codextool

OpenAI’s coding tool/agent used for software development workflows. It matters for PMs as a replacement or alternative in enterprise coding adoption.

AI agentsconcept

Autonomous or semi-autonomous AI systems that can plan and take actions across tools and workflows. This is a core AI PM concept central to product design and evaluation.

Claude Fable 5tool

A Claude model variant being updated with stronger biology safeguards to reduce false positives while still routing dual-use biology requests to higher-safety fallback behavior. Relevant for PMs considering safety tradeoffs and product-surface-specific policy tuning.

vibe-codingconcept

An AI-native development approach where builders use AI tools to rapidly create software. The newsletter treats it as a growth and product-building methodology.

RAGconcept

RAG is a retrieval-based pattern that injects external context into prompts to improve model responses. The newsletter presents it as often outperforming fine-tuning for practical product work.

Kimi K3tool

A 2.8T-parameter open-weight model described as frontier-level by the speaker in the newsletter. It is notable for strong quality and deployment on Nebius Token Factory.

coding agentsconcept

Agents used to write, review, and iterate on code as part of software development workflows. The newsletter frames them as shifting developers toward specification, architecture, and evaluation work.

Fireworks AIcompany

An AI infrastructure company focused on model serving and specialized intelligence. The newsletter mentions its Forge ’26 conference in the context of engineering skills and model operations.

Xcompany

Social platform referenced as a source of examples, discussion, and scraping/monetization concerns. In this newsletter it is part of the agent workflow stack and content source.

Gemini CLItool

Google’s command-line interface for working with Gemini in developer workflows. It is mentioned as a compatible tool alongside agent skills in antigravity.

PixVersetool

A video creation platform with CLI and API access. The newsletter highlights PixVerse's command-line workflow for generating video from prompts and its newer v6 headless engine.

PixVerse AI v6tool

A versioned PixVerse release focused on headless prompt-to-video automation. The newsletter highlights its cinematic realism engine and CLI-based workflow for generating videos programmatically.

Cinematic Realism Enginetool

A headless prompt-to-video engine focused on realism, multi-shot sequencing, and dynamic camera motion. It is framed as the core capability behind PixVerse AI v6's CLI workflow.

Large Memory Modelsconcept

A memory architecture that mimics human memory instead of relying on RAG or vector search. For PMs, it suggests alternative approaches to long-context recall and personalization.

Stay updated on Santiago

Get curated AI PM insights delivered daily — covering this and 1,000+ other sources.

Subscribe Free