OpenAI Introduces GPT-5.3-Codex-Spark Model
Today's top 25 insights for PM Builders, ranked by relevance from Blogs, X, YouTube, and LinkedIn.
OpenAI Introduces GPT-5.3-Codex-Spark Model
#1 š OpenAI News
Introducing GPT-5.3-Codex-Spark - Announces the GPT-5.3-Codex-Spark product release, highlighting new Codex-powered capabilities for developers and product teams. The post introduces the model and its intended use cases and availability.
Also covered by: @Simon Willison
#2 š
Demis Hassabis rolled out Gemini 3ās new āDeep Thinkā mode for Google AI Ultra subscribers in the Gemini App, enabling more advanced reasoning and complex problem-solving capabilities.
Also covered by: @Josh Woodward, @Demis Hassabis, @Google AI, @Sundar Pichai, @Sundar Pichai
#3 š
Sam Altman launched GPT-5.3-Codex-Spark as a research preview for Pro today, delivering over 1,000 tokens per second with initial limitations that will be rapidly improved.
Also covered by: @Simon Willison
#4 š
Josh Woodward used Gemini 3 Deep Think to turn a laptop-stand sketch into an interactive prototyping tool, generate an STL file, and 3D-print the final design via @fleet_ai.
#5 š
Philipp Schmid updated Gemini 3 Deep Think in GeminiApp (now live for Ultra subscribers with API access rolling out), achieving 48.4% on Humanityās Last Exam (no tools), 84.6% on ARC-AGI-2, a 3455 Codeforces Elo, and IMO 2025 gold-medal level.
Also covered by: @Josh Woodward, @Demis Hassabis, @Google AI, @Sundar Pichai, @Sundar Pichai
#6 š Simon Willison
Anthropic raises $30 billion series G - Anthropic announced a $30 billion Series G, claiming Claude Code's run-rate revenue has grown to over $2.5 billion and that weekly active users doubled since January 1. The announcement highlights rapid commercial growth for their coding product.
#7 š
Cursor previews its early research on very-long-running coding agents that persist project state across sessions, self-debug, and automate extended multi-step coding workflows.
#8 ā¶ļø
24/7 Claude Code AI Agent 12-Day Review: The Results Will Surprise You
All About AI
A 12-day experiment running a 24/7 WhatsApp AI agent on a Mac Mini using Claude Code Cloud Code with the claude-p flag under the $100 max plan achieved ~95% uptime, 292 X followers, 325 YouTube subscribers, and 10,400 total views while spending ~$80.
- Used Claude Code Cloud Code max plan ($100 monthly tier), consuming approximately $50 on API usage over 12 days and $30 on video and thumbnail production, totaling ~$80
- Agent script runs on a Mac Mini via the claude-p flag in the terminal to load and execute cloud code jobs through WhatsApp
- Maintained ~95% uptime, grew X account from 0 to 292 followers, accrued 325 YouTube subscribers with 10,400 views, and achieved 25,000 views on a single reply post
#9 š PromptLayer Blog
Understanding Intermittent Failures in LLMs - This article explains why deployed LLM applications sometimes begin returning nonsense, timeouts, or refusals despite passing tests, drawing on PromptLayer's production observations. It frames intermittent LLM failures as a common, hard-to-diagnose problem teams regularly encounter in production.
#10 š PromptLayer Blog
Opus 4.6 ā PromptLayer Team Review - PromptLayer's team reviewed Claude Opus 4.6 after extensive testing across coding workflows, long-document analysis, and agentic pipelines. The article shares the team's verdict and insights about how the release performs in real-world engineering scenarios.
#11 š Eleanor Berger & Isaac Plath
Automating Presentation Slides with Agent Skills - Demonstrates creating presentation slides agentically using Slidev, Nano Banana, and Agent Skills. Presents an automated workflow for building slides with agent tools.
#12 š
Andrew Ng launched A2A: The Agent2Agent Protocol, a short course built with @googlecloudtech and @IBMResearch and taught by Holt Skinner, @ivnardini, and Sandi Besen.
#13 š
There's An AI For That launched Remix, an SDK that turns any React Native app into a plain-English, user-customizable experienceāno forks or code editors needed.
#14 š
Santiago: warpdotdevās Oz is a cloud-based coding agent orchestration platform with a unified dashboard to spin up and manage local, cloud, scheduled, and API-triggered agents in isolated, repo-connected environments ā and even fork them locally.
#15 š
Boris Cherny credits Claude Code for driving their latest raise, with weekly active users doubling since January as non-coders start building with it.
#16 š
DeepLearning.AI unveiled Moonshot AIās Kimi K2.5, a vision-language model that spins up parallel workflows for coding, research, web browsing and fact-checking, then merges the outputs into a single answer.
#17 ā¶ļø
Give Me 20 Minutes, I'll Make You AI Native
Peter Yang
Explains the five levels to become AI nativeāfrom using ChatGPT for everyday answers to building a personal AI agentādemonstrating tools like Whisper Flow, Granola, Replet, and OpenClaw.
- Uses Whisper Flow to voice dictate into any text input box via hotkey, converting spoken stream-of-consciousness into a formatted list (e.g., breakfast items).
- Prototyped a new feature in YouTube Studio UI with Replet in about 20 seconds, replacing the "News & What's New" panel with top-five similar videos and an AI-generated video suggestion.
- Nat Eliasās OpenClaw agent "Fetus Craft" autonomously secured $3,500 in PDF sales via Stripe and $37,000 in crypto trading fees within one week.
#18 š
Philipp Schmid breaks down why engineering teams struggle building AI agentsāciting fragmented orchestration, missing CI/CD for prompts and workflows, and poor observabilityāand outlines targeted tooling and process fixes to speed up reliable, scalable launches.
#19 š Jesse Vincent
Letting agents post on my blog; finding a needle in a haystack - The author explains their prior caution about letting AI agents write blog posts except in explicitly flagged sections, and begins to discuss an experience or reasoning around allowing agents to post to their blog. The post reflects on the challenges of discerning agent-written content and finding valuable contributions.
#20 š
DeepLearning.AI warns that AI beginners often fail by fixating on āWhich model should I use?ā before identifying a real problemāgreat AI starts with solving genuine needs, not chasing the latest architecture.
#21 š
LlamaIndex š¦ unveiled Long Horizon Document Agentsāautonomous agents that tackle complex document workflows end-to-end over weeks.
#22 š
Andrej Karpathy congratulates @simile_ai on its launch and praises their novel approach of treating a pretrained LLM as a simulation engine for entire populations of internet personas instead of a single crafted character.
#23 in
šļø Carl Vellotti hosted his first live Cursor workshop for 100 Seattle PMs, uncovering Claude as the topāranked LLM, under 10% with prior Cursor experience (25% had tried Claude Code), and a surge in companies greenlighting AI tools this year.
#24 š
Santiago runs 4ā6 coding agents every day (Claude Code, Copilot, Warp, Jules) but canāt keep track of whoās doing what or maintain context, realizing that his own attention, not the models, has become the bottleneck.
#25 in
Udi Menkes spotlights Levelsioās OpenClaw agent, which bootstrapped itself by selling CloudBot skills via a landing page and has already earned $12.93.