DeepAgents introduces dynamic subagents for modular AI workflows
Today's top 23 insights for PM Builders, ranked by relevance from Blogs, X, and YouTube.
DeepAgents introduces dynamic subagents for modular AI workflows
#1 đ Anthropic News
More details on Fable 5âs cyber safeguards and our jailbreak framework - Claude Fable 5 has been reâdeployed and is available globally, and Anthropic trained safety classifiers to sort cybersecurity requests into four categoriesâProhibited (blocked), Highârisk dual use (blocked), Lowârisk dual use (monitored/sometimes blocked within a larger âsafety marginâ), and Benign (allowed with monitoring)âwith Fable 5âs safety margin larger than prior models. Anthropic also published an early draft jailbreak severity framework with Glasswing, launched a HackerOne program and feedback channel (cyber-safeguards@anthropic.com) for reporting cyber jailbreaks, and enumerated prohibited actions such as ransomware, cyberâphysical sabotage, AV/EDR bypass, commandâandâcontrol, data exfiltration, malware development/propagation, and internet backbone attacks.
Also covered by: @Claude
#2 đ Claude Code Blog
Giving admins more visibility and control over Claude spend - Announces new admin features that give organizations greater visibility into Claude usage and spending, along with controls to manage costs and governance. The update targets product administrators who need reporting and spend management tools.
#3 đ
xAI has installed Grok Build into Railway sandboxes, letting PM Builders instantly spin up and test AI-powered workflows in their dev environment.
#4 đ
Thariq confirms Fable will exit subscription plans on July 7 but will be reinstated as a standard feature as soon as capacity allows, per the original blog post.
Also covered by: @Claude
#5 đ
Harrison Chase introduced dynamic programmatic subagents in DeepAgents, letting developers spin up on-the-fly subagentsâeach with its own memory, tools, and promptsâto modularize and scale complex AI workflows.
#6 đ
Santiago demoed an agent built on the Linux Foundation-governed x402 open protocol (by Coinbase) that autonomously discovers, pays (in USDC on Base via HTTP 402), and runs Apify Store Actors with no API keys, accounts, or manual billingâenabling true pay-as-you-go agent workf...
#7 đ
LlamaIndex đŚ built an email-processing assistant using LiteParse inside a flueai agent, with Resend webhooks and tursodatabase for persistence, that fetches attachments (including PDFs), parses them, summarizes messages, and drafts replies.
#8 đ
Philipp Schmid built Gemini Omni Flash in just 12 lines using the Interactions API to let you conversationally edit video lightingâupload a clip, request âMake it day time,â and get back a video with shifted shadows and a new sky.
#9 đ
Guillermo Rauch announces AI Gateway Rulesâa CDN-style feature to dynamically reroute or deny AI model traffic on the fly (e.g., rewrite anthropic/claude-fable-5 â anthropic/claude-opus-5)âso you can handle sudden model retirements without redeployment.
#10 đ Ampcode Chronicle
Putting an Agent in an Orb - Amp runs orbs on Debian 12 with preinstalled dev tooling (gh, amp, PostgreSQL, Redis, tmux, ffmpeg, ImageMagick, ripgrep, Bun/Node, npm/pnpm, agent-browser), and a 428-line .agents/setup script that starts PostgreSQL (tuned for speed: fsync=off, synchronous_commit=off, autovacuum=off), creates the amp user/database, seeds test users, installs mise and repo-managed Node/pnpm, runs pnpm install --frozen-lockfile, installs Pillow and sqlite3, writes AGENTS.md guidance, and snapshots the orb for up to 24 hours. The repo also provides a dev-server skill with ensure-dev-server.sh (which ensures .env.local and secrets AMP_API_KEY_SECRET, THREAD_ACTORS_SERVICE_SECRET, THREAD_ACTORS_WEBSOCKET_JWT_SECRET, and can reuse/restart/start the server), writes .amp/dev-ports.json, and exposes /__dev endpoints (including /__dev/log-me-in/
#11 đ Ampcode Chronicle
Read Bigger Threads - Amp rewrote read_thread into a dedicated subagent (now running GLM 5.2 instead of Gemini 3.5 Flash) that searches long threads, verifies whether edits actually succeeded, and explicitly checks for newer messages that revise, supersede, or revert earlier hits. The change was driven by compaction making threads hugeâone thread would be ~21 million tokens without compaction and has been compacted over 68 timesâand the subagent uses compactions for orientation but inspects original messages when exact wording, code, chronology, or verification matter.
#12 đ Simon Willison
llm-coding-agent 0.1a0 - Simon shipped an early alpha of a new Python coding agent built on his LLM/agent framework, created via Claude Fable 5 and released to PyPI and GitHub. The post describes the prompts, spec, README, available tools, and a demo transcript showing the agent generating a simple Swift CLI ASCII clock.
#13 đ Simon Willison
Using DSPy to evaluate and improve Datasette Agent's SQL system prompts - Simon experimented with dspy to evaluate and improve the system prompts used by Datasette Agent for read-only SQL queries, firing off an asynchronous research task using Claude Fable 5. The tool suggested testing with smaller models and highlighted prompt issues like missing column names in schema listings, which can cause guessing and retry loops.
#14 âśď¸
160,000+ Cloned These 3 FREE AI Employees: Here's How (GitHub Claude Skills)
Helena Liu
Installing three free GitHub AI agent repositoriesâLLM Console skill, Last30Days, and G-Stackâinto Claude CoWork desktop via clone commands to enable five debating advisors, comprehensive 30-day sentiment analysis, and a Y Combinatorâstyle AI development team through slash commands.
- LLM Console skill (based on Karpathyâs method) installs five advisorsâContrarian, First Principle Thinker, Expansionist, Outsider, Executorâvia git clone into Claude CoWork and is activated with âconsole thisâ for strategic business debates.
- The Last30Days repository, with over 46,000 GitHub stars, uses the slash command â/last30days <topic>â to scan Reddit, X, YouTube, TikTok, Instagram, and HackerNews and returns summarized sentiment with exact URLs in about five minutes.
- G-Stack by Garry Tan (115,000+ GitHub stars) installs via a single clone command into Claude CoWork to provide roles like CEO, engineering manager, senior designer, QA lead, and tester, accessible through dedicated slash commands such as â/officehourâ.
#15 đ
Boris Cherny uses Claude to auto-generate artifactsâtables, diagrams, color-coded charts and PR overviewsâto streamline architecture/design option comparisons, session result visualization, data analysis and complex code reviews.
#16 đ
Lenny Rachitsky relays OpenAIâs Codex app lead caution that eliminating dedicated product roles to make âeveryone a builderâ is a mistake, because product management is a distinct discipline with unique best practices and skills engineers often overlook.
#17 đ
Aravind Srinivas argues that personal robots are the killer app for running AI models on local hardwareâpeople wonât stream home data to servers, and such devices can double as token faucets for digital tasks.
#18 đ
Lenny Rachitsky asked OpenAI Codex lead @ajambrosino why AI âsucks at design.â He says labs prioritize codingâbecause itâs easier to grade and accelerates AI researchâwhile real design demands cultural insight and novelty, so models default to familiar patterns (e.g.
#19 đ
Julien Chaumond â Co-founder and CTO @huggingface notes that even Palantir now publicly supports open models, underscoring a significant industry shift toward AI openness.
#20 đ
clem đ¤ â Co-founder & CEO @HuggingFace Rampart from @ndstudio & the @WhiteHouse is now the #1 trending token classification model on Hugging Face, marking a shift as public organizations build and own their own model weights instead of renting them.
#21 âśď¸
Claude Fable 5 Is Finally Back: 5 Must-Try Use Cases Before July 7
Peter Yang
Demonstrates five specific high-leverage use cases for Claude Fable 5 on a Claude subscriptionâvia the Claude Code interface with high-effort prompts and API integrationsâbefore its July 7 availability cutoff at 50% weekly usage.
- Claude Fable 5 is available on Claude subscriptions until July 7 with a 50% weekly usage cap, after which it switches to pay-as-you-go API credits.
- Fable 5 audited a vibe-coded fitness app via five parallel agents, running all unit tests and identifying over 12 major bugs including a user-data leak on involuntary sign-out.
- Fable 5 drafted a detailed HTML plan for a new nutrition-tracking tabâcomplete with UI mockups, Supabase schema, and USDA FDC API integrationâusing the Lavish editor plugin.
#22 âśď¸
Fable 5 vs GPT 5.6 Sol: The Early Results
AI Explained
GPT 5.6 Soul is benchmarked against Anthropicâs Fable 5 (Mythos 5) on Terminal Bench 2.1, Healthbench Professional, and Exploit Bench, with GPT 5.6 Soul scoring almost 92% vs 88% on Terminal Bench, 60.5% (64% length-adjusted) vs 66.0% on Healthbench, â76% vs â78% on Exploit Bench, and costing half the API input price and just over half the output price of Fable 5.
- GPT 5.6 Soul on Terminal Bench 2.1 ultra mode scored almost 92% compared to Mythos 5âs 88%.
- On Healthbench Professional, Mythos 5 scored 66.0% raw horsepower versus GPT 5.6 Soulâs 60.5% (64% length-adjusted).
- GPT 5.6 Soul used approximately 120,000â130,000 output tokens on Exploit Bench, compared to 350,000 tokens for Mythos 5.
#23 đ
Harrison Chase unveils LangSmith Engine, an agent launched last month that hunts through your AI agentsâ failures, prioritizes issues, and auto-drafts fixes using sandboxed environments and subagents.