Google Launches Gemini 3.1 Pro

Today's top 25 insights for PM Builders, ranked by relevance from X, Blogs, YouTube, and LinkedIn.

Google Launches Gemini 3.1 Pro

#1 𝕏

Google AI launched Gemini 3.1 Pro for tackling complex workflows, Photoshoot in Pomelli for studio-quality product visuals, and Lyria 3 to turn photos/text into dynamic music.

Also covered by: @Philipp Schmid, @Jason Zhou, @Demis Hassabis

#2 𝕏

Hugging Face welcomes GGML, integrating its lightweight inference library to accelerate on-device ML deployments.

#3 📝 Simon Willison

We’ve made GPT-5.3-Codex-Spark about 30% faster - Thibault Sottiaux (OpenAI) reports a ~30% speed improvement to GPT-5.3-Codex-Spark, which is now serving at over 1200 tokens per second. The note is shared as a short tweet quoted on Simon Willison's weblog.

#4 📝 Simon Willison

Taalas serves Llama 3.1 8B at 17,000 tokens/second - Canadian hardware startup Taalas announced a custom hardware implementation of Llama 3.1 8B running at 17,000 tokens/second using aggressive quantization techniques. Simon notes the demo is so fast it resembles a screenshot and links to chatjimmy.ai to try it.

#5 𝕏

Qwen has launched the Qwen3-Coder-Next API on Alibaba Cloud Model Studio and integrated it into the Coding Plan, giving teams scalable, cost-effective coding endpoints.

#6 𝕏

Teresa Torres shows that when ShowMe restyled its AI sales agent interface to mirror Google Meet/Zoom calls, prospects began interacting with the bot like a human colleague—and conversions surged.

#7 𝕏

LlamaIndex 🦙 launched a hands-on LlamaCloud demo using Agent Workflows (LlamaParse Agentic tier + SQLite) to parse receipt photos and aggregate monthly spending. It then taps Gemini 3.1 Pro to analyze trends and generate actionable financial tips.

#8 𝕏

Claude: TARA, the “Keep Thinking” Prize–winning pipeline by Kyeyune Kazibwe, turns dashcam road footage into detailed economic appraisals and infrastructure investment recommendations. It was validated on an actual road under construction in Uganda.

#9 𝕏

Claude showcases Asep Bagja Priandana’s Opus 4.6 Conductr, a four-track generative band that follows your MIDI chords in real time and runs on a C/WASM engine at ~15 ms latency.

#10 ▶️

Can a Claude Code AI Agent CRUSH The Predictions Market? Let's find out

All About AI

A Claude Code AI Agent with the 'Poly Market/skill MD' skill and cloud code Opus automates browser-based 5-minute Bitcoin up/down trades on Poly Market, generating a 960% return on a $1 bet.

  • Loaded the 'Poly Market/skill MD' skill in Claude Code and used cloud code Opus to control the browser for Poly Market navigation.
  • Placed front-loaded bets into the next 5-minute window based on seven signals: price versus target, Binance websocket price versus target, sidebar consensus, momentum, short trend, crowd positioning, and sidebar shift direction.
  • Recorded a $9 profit on a $1 'up' bet (960% gain) and claimed a total of $37 winnings added to the account.

#11 𝕏

Peter Yang shares a 5-step AI-powered prototyping workflow: use Google AI Studio to turn a screenshot into an interactive base template, co-build and refine new features with AI, then gather feedback from designers and real users.

#12 ▶️

Gemini 3.1 + New AI Studio Is Here: Full Prototyping Tutorial in 18 Minutes

Peter Yang

Google Gemini 3.1 and Google AI Studio's new full-stack update replicate the existing AI Studio UI and simplify it through a five-step prototype-first workflow, using custom Gemini prompts to produce a redesigned interface in roughly 141 seconds.

  • Google Gemini 3.1 and Google AI Studio's full-stack update support in-tool servers, databases, and multiplayer features.
  • Replicating the existing AI Studio UI as a base template took approximately five minutes before saving it as "AI Studio template."
  • Submitting a custom simplification prompt to a remixed AI Studio template generated the redesigned interface in about 141 seconds.

#13 📝 PromptLayer Blog

How to Install OpenClaw: Step-by-Step Guide (formerly ClawDBot / Moltbot) - A hands-on installation guide for OpenClaw, a popular always-on assistant project in the agentic AI community, walking readers through setup and explaining what OpenClaw does.

#14 📝 PromptLayer Blog

Claude Opus 4.1 (20250805 Thinking 16k): What the 'Thinking 16k' Label Actually Means for Your Workflows - Explains the naming convention for Claude Opus 4.1 and clarifies that the long slug refers to a reasoning-budget configuration of Anthropic's flagship model rather than a separate model. Helps readers understand implications for workflow and model selection.

#15 in

Greg Isenberg reveals that he almost shelved OpenClaw before realizing he needed to narrow its feature set, run rapid user tests, and focus distribution on niche communities. Those pivots revived the stalled product and set it on a growth path.

#16 ▶️

Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI

AI Explained

Demonstrates Gemini 3.1 Pro’s performance across diverse benchmarks with 77.1% on ARC AGI 2, 79.6% on a private Simple Bench test, and a reduction in fine-tuning runtime from 300 seconds to 47 seconds.

  • Gemini 3.1 Pro scored 77.1% on the ARC AGI 2 puzzle benchmark, outperforming Claude Opus 4.6’s ~69% and GPT 5.2 extra-high’s ~50%
  • On the private Simple Bench test of trick questions and common-sense reasoning, Gemini 3.1 Pro achieved 79.6%, matching the human average baseline from nine participants
  • Fine-tuning runtime using Gemini 3.1 Pro decreased from 300 seconds to 47 seconds, surpassing the 94-second human reference solution

Also covered by: @Philipp Schmid, @Jason Zhou, @Demis Hassabis

#17 𝕏

Santiago unveils Surething, an autonomous chat-based agent that in 5 seconds connects to 1,000+ apps (email, Slack, Zoom, etc.) with zero setup and lets you build and schedule no-code automations by simply asking.

#18 𝕏

Google AI launched Photoshoot, a free Google Labs tool that auto-generates on-brand, campaign-ready product visuals in three modes—templates, edits to existing images, or custom-from-scratch creations.

#19 𝕏

DeepLearning.AI urges letting real users break your AI prototypes early to uncover hidden flaws. Test small, learn fast, and fix early to accelerate improvement.

#20 𝕏

DeepLearning.AI: The nonprofit AI Verification and Research Institute (Averi) launched standards for independent audits of AI systems—assessing misuse, data leaks, and harmful behavior—to embed safety reviews into development cycles and build public trust.

#21 𝕏

Logan Kilpatrick shares a YouTube deep dive into music-making, unpacking the software tools and creative workflows behind song production.

#22 𝕏

Logan Kilpatrick had a fun chat with the Lyria 3 team about launching their newest music model in the Gemini App this week.

#23 𝕏

Sebastian Raschka lists February’s AI model launches—Moonshot AI’s Kimi K2.5, z.AI’s GLM 5, MiniMax M2.5, ByteDance’s Seed-2.0, Nanbeige 4.1 3B, Qwen 3.5, and Cohere’s Tiny Aya. He also flags anticipation for DeepSeek V4 soon.

#24 𝕏

Andrej Karpathy notes that while most LLM heavy lifting happens on provider clusters and PicoClaw can run on under $10 of hardware, a $599 M4 Mac mini still offers outstanding value and headroom for local model experimentation.

#25 𝕏

Andrej Karpathy bought a Mac mini to tinker with OpenClaw but warns it’s a “wild west” security nightmare—exposed instances, RCE flaws, supply-chain poisoning and malicious registry skills abound.

Get tomorrow's brief first

Join AI product managers receiving the latest brief before it reaches the public archive.

Subscribe free