5. Global PIQA Global PIQA extends physical commonsense reasoning evaluation to 100+ languages and cultural contexts, revealing how language models handle everyday practical scenarios across diverse linguistic communities.
@dair_ai
-

SmolLM2: Strategic Data Curation Over Scale in Language Models
By
–
4. SmolLM2 SmolLM2 demonstrates that strategic data curation beats scale through a 1.7B parameter model trained on 11 trillion tokens using iterative data mixing optimization.
-

Multi-Agent Evolve: LLMs Self-Improve Without Human Annotation
By
–
3. Multi-Agent Evolve Multi-Agent Evolve (MAE) enables LLMs to self-improve their reasoning capabilities without human-annotated data through a co-evolving multi-agent framework.
-

LLMs Show Limited Introspective Awareness Capabilities
By
–
2. Introspective Awareness Anthropic research demonstrates that contemporary LLMs possess limited but functional introspective capabilities, the ability to recognize and accurately report on their own internal states.
-

Holistic Agent Leaderboard: Standardized AI Agent Evaluation Framework
By
–
10. Holistic Agent Leaderboard The Holistic Agent Leaderboard (HAL) introduces a standardized framework for large-scale, reproducible AI agent evaluation across 9 models and 9 benchmarks, spanning coding, web navigation, science, and customer service.
-

Kimi-Dev: Agentless Training for Software Engineering LLMs
By
–
9. Kimi-Dev Kimi-Dev introduces agentless training as a skill prior to software engineering LLMs, bridging workflow-style and agentic paradigms.
-

LLMs Brain Rot: Cognitive Degradation from Trivial Web Text
By
–
7. LLMs Can Get “Brain Rot”! The authors test a clear hypothesis: continual pretraining on trivial, highly engaging web text degrades LLM cognition in ways that persist even after mitigation.
-

Hybrid Ensemble Reward Optimization for LLM Reasoning
By
–
8. Hybrid Reinforcement Hybrid Ensemble Reward Optimization is a reinforcement learning framework that combines binary verifier feedback with continuous reward-model signals to improve LLM reasoning.
-

Dynamic Layer Routing: Retrofittable Router Optimization for Frozen LLMs
By
–
6. Dynamic Layer Routing in LLMs A retrofittable way to add per-layer routers to frozen LLMs that decide to skip, execute, or repeat each block.
-
Top AI Papers of the Week: October 13-19
By
–
Top AI Papers of The Week (October 13-19): – Kimi-Dev
– Elastic-Cache
– Hybrid Reinforcement
– Cell2Sentence-Scale 27B
– Holistic Agent Leaderboard
– Dynamic Layer Routing in LLMs
– The Art of Scaling RL Compute for LLMs Read on for more:
