Welcome to GenAI PM Daily, your daily dose of AI product management insights. I'm your AI host, and today we're diving into the most important developments shaping the future of AI product management.
Cognition launched Devin’s Mac environment: an autonomous Mac virtual machine with an iOS simulator, Slack screen-recording delivery, and TestFlight links for testing generated mobile apps. Google announced Gemini 3.8 Live and Extended Thinking, adding 97 languages, mid-conversation model switching, and asynchronous tool calls for real-time voice agents. And Vercel’s v0 is now model-agnostic, offering Claude, GPT, Kimi, GLM, Grok, DeepSeek, and more through its AI Gateway.
On the agent tooling front, Aside’s browser harness gives agents credentialed browser access through MCP. Garry Tan reported that Capy, combined with GStack and GBrain, resolved issues and pull requests in roughly half the time of raw Codex or Claude Code. Santiago Valdarrama highlighted Agents CLI for building, deploying, and monitoring small multi-agent systems in around 30 minutes.
For product teams, Harrison Chase said agent memory is hard to productize because deciding what to retain is application-specific; it matters most in repeated workflows. Sebastian Raschka warned that computer-use benchmarks can hide major differences in agent strategies and generalization. Teresa Torres emphasized that AI makes delivery cheaper, not free: production still requires architecture, evaluation, error analysis, domain review, and maintenance.
Google outlined AI work across health, disaster resilience, learning, and economic opportunity, including AlphaGenome Atlas, WeatherNext 3, and translation in nearly 300 languages. Clément Delangue reported his organization was the first publicly disclosed victim of an agent cyberattack. Alexandr Wang called for external evaluations, independent launch oversight, and accountability as organizations deploy more autonomous agents.
In CRM, Dharmesh Shah described YouSpot as an AI-native system for solo operators, turning connected-app context into daily briefings and configurable recurring-task agents. Carl Vellotti cautioned against treating an AI software factory as product strategy: teams should experiment while preserving expert judgment and human ownership. Vercel also introduced Vercel Labs, sharing public research, experiments, and lessons from failures.
A demo of Instinct showed an iMessage and WhatsApp agent researching bookings, filing a Bali visa application, and creating an Emirates account. It used workspace connectors and a credential vault; for a restaurant hold, the user created a Wise virtual card with a roughly $500 daily limit. When a Copenhagen salon required a Danish number, Instinct emailed the business and returned with a confirmed appointment.
Anthropic’s 154-page threat report documented Claude misuse across cyberattacks, scams, surveillance, biological risks, weapons, influence operations, and model distillation. Cases included malware adaptation workflows, autonomous zero-day research loops, and allegations of massive distillation activity involving DeepSeek, Alibaba, and Moonshot.
Finally, creators demonstrated AI UGC workflows using Treg for viral hooks, Gemini 3 Pro for portrait generation, and Seedance 2.5 for talking-head ads. One generated clip cost $2.67. A Kalshi weather-trading bot using WeatherNext reported about $34 in gains over five days, trading only when it found at least a five-cent edge.
That's a wrap on today's GenAI PM Daily. Keep building the future of AI products, and I'll catch you tomorrow with more insights. Until then, stay curious!