"Representational Curvature Modulates Behavioral Uncertainty in LLMs" Most interpretability work studies what features live inside LLMs. But this paper studies something deeper, which is the geometry of the whole evolving representation. The key result is that representational
MACHINE LEARNING
-
Monetizing machine learning models with XGBoost
By
–
Is your XGBoost model making money for you while you sleep? @trainxgb @tabul_ai
-

Hardware limitations on sparse neural activation in LLMs
By
–
The human brain is incredibly efficient because it only activates the specific neurons needed for a thought. Modern LLMs naturally try to do this too (> 95% of neurons in feedforward layers stay silent for any given word), but our hardware punishes them for it. One of the most
-
Optimiser les LLM avec la sparsité adaptée au GPU
By
–
How do we make LLMs faster and lighter? Don’t force the GPU to adapt to sparsity. Reshape the sparsity to fit the GPU! ⚡️
— Sakana AI (@SakanaAILabs) 8 mai 2026
Excited to share our new #ICML2026 paper in collaboration with @NVIDIA: "Sparser, Faster, Lighter Transformer Language Models". This work introduces new… pic.twitter.com/ehByWHIh6IHow do we make LLMs faster and lighter? Don’t force the GPU to adapt to sparsity. Reshape the sparsity to fit the GPU! Excited to share our new #ICML2026 paper in collaboration with @NVIDIA
: "Sparser, Faster, Lighter Transformer Language Models". This work introduces new -
Internal Manual for Building AI Agent Skills Published
By
–
We've published our internal manual for building agent skills. Skills require a new way of thinking for developers.
-

Addressing High Costs in AI Coding via Open-Source Models
By
–
ai coding is getting expensive use more open models!
-

SkillOS: Skill Curation for Self-Evolving Agents
By
–
SkillOS Learning Skill Curation for Self-Evolving Agents paper: https://
huggingface.co/papers/2605.06
614
… -

Continuous-Time Distribution Matching for Few-Step Diffusion Distillation
By
–
Continuous-Time Distribution Matching for Few-Step Diffusion Distillation paper: https://
huggingface.co/papers/2605.06
376
… -

Apple introduces TIDE: Every Layer Knows the Token Beneath the Context
By
–
Apple presents TIDE Every Layer Knows the Token Beneath the Context paper: https://
huggingface.co/papers/2605.06
216
…
