New AI Models Launch: GPT-4.5, Microsoft's Phi-4 and IBM's Granite 3.2 Unveiled
Today's curated insights on AI product management, selected by our AI agent from 1000+ updates across 50+ expert sources.
New AI Models Launch: GPT-4.5, Microsoft's Phi-4 and IBM's Granite 3.2 Unveiled
From Twitter
Here’s a categorized summary of the key discussions:
New AI Model Releases & Updates
- Sam Altman announced GPT-4.5, highlighting its more natural conversational abilities but noting it’s “not a reasoning model.” Currently available to Pro users due to GPU constraints.
- Microsoft released Phi-4 models including a 3.8B Instruct version and 5.6B Multimodal version supporting audio and images in 23 languages
- IBM unveiled the Granite 3.2 model family featuring compact reasoning and vision-language models
- Inception Labs launched Mercury, a diffusion AI family generating text ~10x faster than traditional LLMs
Product & Platform Updates
- Anthropic introduced hierarchical summarization to help identify usage patterns and misuse in Claude’s computer use capability
- Poe announced Poe Apps, enabling users to create visual UIs using 100+ models
- LangChain released v0.3 with prebuilt agents and new capabilities
- Meta introduced Aria Gen 2 glasses for research in machine perception and contextual AI
AI Development & Best Practices
- Andrew Ng shared insights on the Voice Stack and best practices for voice-based applications, including handling latency and response quality
- MUFG Bank reported a 10x boost in sales efficiency using LangChain for automated data extraction and presentation generation
- LlamaIndex launched LlamaExtract for structured data extraction from unstructured documents
Industry Milestones
- Waymo reported serving 200k+ paid trips weekly across LA, Phoenix and SF, representing 20x growth in two years
- Hume AI released Octave, a text-to-speech LLM understanding emotional context
- Amazon debuted Alexa+, a next-gen AI assistant with enhanced personalization and memory
Memes & Humor
- Sam Altman joked about social apps: “ok fine maybe we’ll do a social app”
- Andrej Karpathy pointed out an amusing inconsistency in GPT-4.5’s understanding of model speed
- Lex Fridman tweeted about anticipating the Epstein list release with popcorn emojis
From Reddit
Theme 1. Coding AI Showdown: Claude 3.7 vs. Grok-3
-
I tested Claude 3.7 Sonnet against Grok-3 and o3-mini-high on coding tasks. Here’s what I found out (Score: 152, Comments: 38): Claude 3.7 Sonnet outperformed Grok-3 and o3-mini-high in coding tasks, excelling in projects like a Minecraft game and a markdown editor, while o3-mini-high struggled with complex tasks and only excelled in a code diff viewer. Read more here.
- Claude 3.7 Sonnet is favored for its coding capabilities without the “thinking” model, while Grok 3 is anticipated to improve with a forthcoming API release, and o3-mini-high stands out in specific tasks but struggles with complexity.