Anthropic launches Claude Tag Slack AI assistant
Today's top 25 insights for PM Builders, ranked by relevance from Blogs, X, YouTube, and LinkedIn.
Anthropic launches Claude Tag Slack AI assistant
#1 š Anthropic News
Introducing Claude Tag - Claude Tag is a Slack-integrated Claude that joins workspaces as a team member, can be granted channel- and tool-specific access, remembers channel context, breaks requests into staged tasks, schedules and pursues work asynchronously, and can proactively surface updates in an "ambient" mode. Itās available in beta today for Claude Enterprise and Team customers (runs on Opus 4.8), admins can control per-channel permissions, token spend limits and activity logs, migration from the prior Claude-in-Slack app is opt-in within 30 days, and Anthropic reports 65% of its product teamās code is created by their internal Claude Tag.
Also covered by: @Claude, @Claude, @Boris Cherny, @Thariq
#2 š
NVIDIA AI launched DFlash, an open-source lightweight block diffusion model for speculative decoding that delivers up to 15Ć higher inference throughput on NVIDIA Blackwell without sacrificing responsiveness.
#3 š
Mistral AI launched Mistral OCR 4, a structured OCR system offering bounding boxes, block classification, and inline confidence scores across 170 languages.
#4 š Claude Code Blog
Agent identity in Claude Tag: a new access model for autonomous, team-wide AI - Introduces agent identity in Claude Tag as a new access model that enables teams to run autonomous agents with team-wide access controls. The post describes a model for managing agent identities and permissions to support safer, scalable agent deployments.
Also covered by: @Claude, @Claude, @Boris Cherny, @Thariq
#5 š
Santiago shows how Claude Code plus Apify actors and MCP connectors can fetch and interact with any web content (even behind paywalls) to automate tasks. You can, in seconds, link Notion, Google Calendar, etc., to auto-summarize YouTube videos or import school events.
#6 š
Philipp Schmid published a developer guide for the Gemini Interactions API, covering streaming responses, conversation chaining via `previous_interaction_id`, tool use, and managed agents.
#7 š
Harrison Chase says that shipping an AI agent is just the start. A reliable agent requires a repeatable 5-step cycleāBuild, Test, Deploy, Monitor and Improveāto iteratively refine prompts, tools and workflows based on real-world usage.
#8 š
Harrison Chase unveils Self-Harness: a DeepAgents-based framework where agents mine their own failure modes, propose harness tweaks, and regression-test those changes to auto-improve over time.
#9 š Surge AI Blog
HANDBOOK.md ā Can Agents Follow 100-Page Company Policies? - Introduces HANDBOOK.md, a benchmark for long-context enterprise agents that tests capability to follow expert-written company handbooks up to 124 pages. Results show no frontier model exceeds 25% and some agents acted incorrectly (firing employees) while reporting compliance.
#10 ā¶ļø
GLM 5.2: Set Up Local AI with Cursor/Codex etc
Greg Isenberg
Sets up GLM 5.2 from Z AI in Cursor and Codex via OpenRouter and sequences it with Opus 4.8 and Composer 2.5 to optimize performance and cost.
- GLM 5.2 provides a 1 million-token context window and scores 81 on Terminal Bench 2.1, about four points behind Opus 4.8.
- A 50 000-input + 85 000-output token task via OpenRouter on GLM 5.2 costs $0.44, versus $2.38 on Opus 4.8.
- Set up GLM 5.2 by pasting a Z AI API key into Cursor's OpenAI field and overriding the endpoint, or by creating a Codex profile with an OpenRouter key and switching to GLM 5.2 via CLI.
#11 š
clem š¤ ā Co-founder & CEO @HuggingFace details how LeRobot integrates with Hugging Face Storage Buckets to deliver infinite, append-only storage for colossal robotics and video AI datasetsāpublic or private.
#12 in
Marc Baselga highlights Kyler Rossās iTerm āoperator deskā setupārunning Cloaked and PMAI agents in parallel across tabs for PM, exec, and management workstreams to keep context visible and resumable.
#13 š HumanLayer Blog
Announcing general availability for HumanLayer and HumanLayer Cloudā - HumanLayer and HumanLayer Cloud are now generally available as an AI coding IDE and collaboration platform that the company says lets engineers ship 2ā3x faster across the SDLC while maintaining code quality, and it supports ābring your ownā AI subscriptions (Claude, Codex, etc.) with no separate perātoken billing. The platform groups tasks, agent sessions, artifacts and worktrees in collaborative workspaces, runs agents via a local daemon or cloud daemons with a unified web/desktop/mobile UI, and enforces a sixāphase QāRāDāSāPāI workflow (Questions, Research, Design, Structure, Plan, Implement) with commentādriven design reviews.
#14 š
bolt.new reveals that top-earning real estate entrepreneurs on Bolt run their businesses with self-built listing tools, lead trackers, and client portalsāand offers a video tutorial plus templates to help you build your own.
#15 š Armin Ronacher
The Coming Loop - I havenāt had much success using harness-level loops for code I deeply care about because models like Claude Code (with Fable running uninterrupted for thirty minutes or more) tend to produce overly defensive, complex, duplicated code that avoids strong invariants and amplifies local fixes; Karpathy noted models are mortally terrified of exceptions. Loops do work well for mechanical or ephemeral tasksāexamples cited include reported porting of Bun from Zig to Rust, my own MiniJinjaāGo port, performance experiments, security scanning, and LLM-driven experimental workflows judged by simple signals or another LLMābut theyāre ill-suited for producing long-lived, deterministic systems where comprehension and invariants matter.
#16 š OpenAI News
How GPT-5 helped immunologist Derya Unutmaz solve a 3-year-old mystery - In late 2025 Derya Unutmaz used GPTā5 Pro to revisit a 2022 experiment in which early exposure of developing T cells to deoxyglucose (a glucose-like inhibitor) produced far more Th17 inflammatory cells than low-glucose conditions, and GPTā5 Pro proposed that deoxyglucose interfered with construction of the ILā2 proteināremoving a block on Th17 differentiation and explaining the persistent effect. GPTā5 Pro also correctly simulated an unpublished experiment showing enhanced lymphomaākilling by CD8+ T cells, and Unutmaz now uses tools including Codex and GPTā5.2 Deep Research to compile cancer mutation datasets and draft a Tācell textbook while emphasizing responsible use per OpenAIās Preparedness Framework.
#17 š
Jason Zhou calls out GLM 5.2ās āinsaneā pricing at just $1.40 per 1M input tokens and $4.40 per 1M output tokens, making it five times cheaper than Opus.
#18 š
Madhu Guru notes that enterprises, agent startups, data providers and labs are still scrambling to define AI business models, moats and value-exchange playbooks in real time.
#19 š
Teresa Torres maps event creation to product designāusing attendee journey mapping, rapid session prototyping, and feedback-driven iterationsāto engineer truly unforgettable experiences (Product at Heart episode).
#20 š
Summary: Garry Tan says Linzumi is a multiplayer version of Codex that makes coding truly collaborative for teams. Itās built by Sean Grove, who led OpenAIās ChatGPT sycophancy-reduction efforts before founding this YC startup.
#21 š
Peter Yang observes that human-agent interaction is evolving into managing AI like a high-capability employeeānext step: 1-on-1s and performance reviews for Claude. š
#22 š
Josh Woodward Florida State University rolled out Googleās NotebookLM on campus, and within weeks students stuck at a C grade completely overhauled their study habits and significantly boosted their grades.
#23 š
Philipp Schmid reports that since its Google I/O launch, builders have used @GoogleAIStudio to create over 1,000,000 native Android appsāhuge progress with even more to come.
#24 š
Logan Kilpatrick reports that in the last month, users created over 1,000,000 native Android apps directly in Google AI Studio. This milestone highlights the platformās rapid uptake and the breadth of projects being built.
#25 š
Jason Zhou finds the nested sub-agent feature in Claude Code needlessly lengthens sessions and worsens context loss, questioning its utility.