As Piotr explains, this is a nasty and unexpected example of @simonw
's "lethal trifecta":
AGI
-

Understanding AI’s Lethal Trifecta Risk Framework
By
–
-
The AGI mania in San Francisco
By
–
The tallest thing in San Francisco right now is the Dirac delta of AGI mania.
-
Gemini 3.1 Pro Doubles Reasoning Capability on ARC-AGI-2
By
–
Introducing Gemini 3.1 Pro 🚀
— Google AI (@GoogleAI) 19 février 2026
3.1 Pro represents a major step forward in core reasoning. It scored 77.1% (more than doubling 3 Pro’s score) on ARC-AGI-2, the benchmark that evaluates a model's ability to solve new logic patterns and work through challenges it hasn’t encountered… pic.twitter.com/8HOPoji7i5Introducing Gemini 3.1 Pro 3.1 Pro represents a major step forward in core reasoning. It scored 77.1% (more than doubling 3 Pro’s score) on ARC-AGI-2, the benchmark that evaluates a model's ability to solve new logic patterns and work through challenges it hasn’t encountered
-

Google drops Gemini 3.1 Pro with 77.1% ARC-AGI-2 score
By
–
BREAKING: Google just dropped Gemini 3.1 Pro. 77.1% on ARC-AGI-2. More than double the reasoning of 3 Pro. This changes everything. Here's what you need to know: First, what is ARC-AGI-2? It's the hardest AI benchmark on the planet. It tests whether a model can solve
-
Terminus 2 Agent Performance Below 5pta Benchmark
By
–
It is also 5pta below Terminus 2, which is meant to be the simplest agent without any tricks
-

Gemini 3.1 Pro Preview Available on Google AI Studio
By
–

BREAKING : Gemini 3.1 Pro Preview is now available on Google AI Studio! "Our latest SOTA reasoning model with unprecedented depth and nuance, and powerful multimodal understanding and coding capabilities"
-

Agent Memory Benchmarks Don’t Predict Real-World Performance
By
–
Agent memory benchmarks are misleading. Scoring well on memory recall doesn't mean an agent can actually use that memory to take correct actions across sessions. Models that achieve near-saturated performance on existing long-context memory benchmarks like LoCoMo perform poorly
-
LLM Intelligence and Human Cognitive Abilities
By
–
Not that odd if you really understand how *human* intelligence works. We’ve known about the existence of the “g-factor” for about a century now. If *all* of the LLM training data is downstream from human cognitive abilities, then it’s intuitively unsurprising that these abilities
-

AI-generated personal MBA curriculum with prompts
By
–
R.I.P Harvard MBA. I built a personal MBA using 12 prompts across Claude and Gemini. It teaches business strategy, growth tactics, and pricing psychology better than any $200K degree. Here's every prompt you can copy & paste:
-

Google’s multi-agent AI framework showcased
By
–
Holy shit… Google just published one of the cleanest demonstrations of real multi-agent intelligence I’ve seen so far. Not another “look, two chatbots are talking” demo. An actual framework for how agents can infer who they’re interacting with and adapt on the fly. The