We're working with @IneffableLabs to co-design the infrastructure for large-scale, reinforcement-learning agents and accelerate discovery across science and industry.
Our engineers have teamed up to explore how to create the training pipeline that will allow agents to discover
MACHINE LEARNING
-

Collaboration to Develop Infrastructure for Large-Scale RL Agents
By
–
-

New research challenges the assumption that multi-agent systems improve LLM reasoning
By
–
Do multi-agent systems make LLM reasoning better? Most AI devs assume that it should. But this new paper shows that this is often not the case. It ran 22,500 deterministic trajectories across GAIA, SWE-bench, and Multi-Challenge with three frontier models. Agents frequently
-

Do Enterprise Systems Need Learned World Models?
By
–
Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
-

Meta-RL with Rubric-guided Policy Decomposition Research
By
–
RubricEM Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
-

Low-commitment attention modification for model training
By
–
Interesting paper. What I like about this is that it is a relatively low-commitment attention modification. I.e., one can use it during most of training, switch back to vanilla attention near the end, and recover roughly the same modeling performance as if full attention had
-

EgoMemReason: A Memory-Driven Benchmark for Egocentric Video Understanding
By
–
EgoMemReason A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Understanding
-

G²RPO-A: A New Training Method for Enhancing LLM Reasoning
By
–
Why can't smaller language models match larger ones on reasoning? Researchers from CUHK Shenzhen, Alibaba Group, and Westlake University introduce G²RPO-A. It adaptively feeds correct reasoning steps into training, dynamically adjusting guidance as the model improves. On math
-

Introduction to Embedded Language Flows Research and Implementation
By
–
ELF: Embedded Language Flows Paper: https://
arxiv.org/abs/2605.10938
Code: https://
github.com/lillian039/ELF Our report: https://
mp.weixin.qq.com/s/7x8w_2Ov-lpS
EPRgqxLZBQ
… -

Kaiming He’s Team Introduces ELF for Continuous Embedding Language Modeling
By
–
Huge! Language models could generate text as smoothly as AI creates images now! Kaiming He's team introduces ELF (Embedded Language Flows). Instead of operating on discrete tokens, ELF stays in a continuous embedding space until the final step, borrowing proven techniques like
-

IndiaAI Data & AI Labs Trains Learners in Foundational AI Skills
By
–
Behind every AI model is quality data, through IndiaAI Data & AI Labs, learners are being trained in this foundational skill that enables everything from chatbots to computer vision systems. 900+ trained learners are forming a critical layer of India’s AI capability building.