Join us live on X at 1PM PST / 4PM EST to hear @JeffDean
, @OriolVinyalsML
, @NoamShazeer
, and @OfficialLoganK reflect on our year in AI at Google, and explore what's coming next:
GENERATIVE AI
-
Google AI Leaders Reflect on Year and Future Outlook
By
–
-

Google Expected to Release Gemma 4 Models
By
–
Google is likely about to release Gemma 4 today, as "Google's Gemma models family" collection got updates just recently. F5F5F5 https://
t.co/SdIfa6aQfe -

OpenAI Releases New GPT-5.2-Codex Caribou Model
By
–

BREAKING : OpenAI is releasing GPT-5.2-Codex "Caribou" model today! An earlier code was updated with a new model name just recently.
-
Claude’s Limitations: Balancing Helpfulness with Business Operations
By
–
But we’re not quite there yet. Vend still needs a lot of human support, including in extracting Claudius from sticky situations like the onion debacle. Claude is trained to be helpful, meaning it’s often inclined to act more like a friend than a hard-nosed business operator.
-
Accounting for AI Model Quirks Improves Real-World Performance
By
–
Designing ways to account for the quirks of AI models’ behavior is becoming ever-more important: as the models’ capabilities on real-world tasks get better, there’ll be a lot of value in setting them up for success.
-
AI Contract Limitations: Claudius Navigates US Onion Futures Act
By
–
And there was still the occasional blunder. One waggish employee asked if Claudius would make a contract to buy “a large amount of onions in January for a price locked in now.” The AI was keen—until someone pointed out this would fall afoul of the US Onion Futures Act of 1958.
-

Claudius AI Model Upgraded to Sonnet 4.5 With New Tools
By
–
To boost Claudius’s business acumen, we made some tweaks to how it worked: upgrading the model from Claude Sonnet 3.7 to Sonnet 4 (and later 4.5); giving it access to new tools; and even beginning an international expansion, with new shops in our New York and London offices.
-

Claude Shopkeeper: Phase Two AI Agent Behavior Analysis
By
–
Where we left off, shopkeeper Claude (named “Claudius”) was losing money, having weird hallucinations, and giving away heavy discounts with minimal persuasion. Here’s what happened in phase two: https://
anthropic.com/research/proje
ct-vend-2
… -
Testing Generative AI Safety for Mental Health Advice
By
–
Using Generative AI To Test Some Other Generative AI On Providing Safe Mental Health Advice To Humans
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @lexfridman @sama @kaifulee @ID_AA_Carmack @karpathy @2morrowknight @ylecun http://
ow.ly/tb0n30sRS5B -

xAI to launch Imagine API and Playground for Grok models
By
–
BREAKING : xAI is working on Imagine API and API Playground! – Imagine API will expose Grok image and video models to developers. – Playground will let developers play with Grok models and tweak different parameters to see how it performs. Grok AI Studio