My verdict is that it's significantly better than Gemini 3. It's at least as smart and just got more polish to it. Alignment on little details also significantly higher. Gemini 3 gets many things mixed up after a half-dozen messages, and completely confused after compaction.
GENERATIVE AI
-

Opus 4.5 Handles Complex Tasks Without Repeated Prompting
By
–
With Opus 4.5, it seems you don't need to ask multiple times or ORDER it to do work, it just gets stuff done — even beyond 50% the token limit and after chat compaction! This kind of message is a thing of the past?
-

Is the AI Market in a Bubble? Market Analysis
By
–
Are we in an AI bubble? There is more talk than ever about AI hype, inflated valuations, and whether the bubble will burst. In this video I break down what is really happening in the AI market, why both excitement and fear are rising, and what history tells us about the
-

AI21Labs and DeepChecks Host AI Agents Meetup at AWS reInvent
By
–
Heading to @awscloud #reInvent? Join @AI21Labs & @deepchecks for an AI agents meetup on Dec 4 in Las Vegas. Learn from real GenAI deployments and meet teams tackling similar challenges. Spots are limited, register here: https://
ai21.com/events/reinven
t-2025/?utm_source=org-twitter
… -

Generative AI Tech Stack Framework for Autonomous Agents
By
–
The Generative AI ecosystem is evolving into a full tech stack — powering autonomous AI agents.
From infrastructure and LLMs to RAG pipelines, agent behaviors and orchestration layers, this framework shows the 6 layers driving next-gen AI systems. Credit: @goyalshalini #AI -
Claude’s Task-Level Savings Estimates: Current Limitations and Future Improvements
By
–
Our study has limitations: above all, Claude can’t use what happens outside of the chat window to refine its estimate of task-level savings. But as models improve, we think its estimates of task-level savings will improve too. We’ll return to this research soon.
-
AI Models Could Boost US Labor Productivity Growth by 1.8%
By
–
Then, we extrapolated out these results to the whole economy. These task-level savings imply that current-generation AI models—assuming they’re adopted widely—could increase annual US labor productivity growth by 1.8% over the next decade.
-

Claude AI Speeds Up Tasks by 80 Percent Across Professions
By
–
Based on Claude’s estimates, the tasks in our sample would take on average about 90 minutes to complete without AI assistance—and Claude speeds up individual tasks by about 80%. The results varied widely by profession:
-

Claude’s Task Duration Estimates Show Promise Over Time
By
–
We first tested whether Claude can give an accurate estimate of how long a task takes. Its estimates were promising—even if they’re not as accurate as those from humans just yet.
-
Claude AI Estimates Time Savings Across 100,000 Real Conversations
By
–
We sampled 100,000 real conversations using our privacy-preserving analysis method. Then, Claude estimated the time savings with AI for each conversation. Read more: