What actually is GBrain? (Y Combinator CEO's personal agent brain) Every agent memory tool you've seen solves a simple problem: store facts, retrieve facts. GBrain solves a different one. It gives your agent a knowledge system that wires itself, enriches itself, and compounds
AI
-
Depth Anything V2: Technical Capabilities and Model Specifications
By
–
-

Technical Analysis of Attention Drift in Speculative Decoding Models
By
–
“Attention Drift: What Autoregressive Speculative Decoding Models Learn” Speculative decoding makes LLM inference faster, but drafters break under small template changes and long context. But why? This paper shows that as the drafter predicts more tokens, its attention drifts
-

Improving LLM Embedding Representations via Mean-Pooling of Generated Tokens
By
–
“The Truth Lies Somewhere in the Middle of the Generated Tokens” LLMs don’t store the meaning of a prompt in one hidden state. As they generate, the meaning gets spread across many token embeddings. So this paper propose a mean-pool over generated token embeddings instead of
-
Mini Shai-Hulud attack threatens AI coding workflows like CI and hooks
By
–
the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc
-
Opus 4.7 2.5x speed at 6x cost surprises user
By
–
i thought this was a joke. Opus 4.7 2.5x speed at 6x cost. what
-

Subquadratic Unveils New SubQ Model Using Subquadratic Sparse Attention
By
–
What if LLMs could read a million tokens without exploding in cost? Subquadratic, an AI research startup, just unveiled SubQ — the first model built on fully subquadratic sparse attention (SSA). Instead of comparing every token pair, SSA routes attention only to the truly
-
Fast Mode for Claude Opus 4.7 Now Available in Cursor
By
–
Fast mode for Claude Opus 4.7 is now available in Cursor! It's 2.5x the speed at 6x the cost. For most tasks, we recommend using the standard speed.
