

Qwen3-Max Thinking is now available on Qwen Chat with 82k tokens worth of thinking budget.

By
–


Qwen3-Max Thinking is now available on Qwen Chat with 82k tokens worth of thinking budget.
By
–
9. Kimi Linear Kimi Linear introduces a hybrid linear attention architecture combining Kimi Delta Attention (KDA) with periodic full attention layers at a 3:1 ratio, achieving superior performance over full attention while reducing KV cache by 75% and delivering 6× faster

By
–
7. Stress-Testing Model Specs This research examines how well large language models adhere to their stated behavioral guidelines by stress-testing AI constitutional specifications through value-tradeoff scenarios.

By
–
5. Global PIQA Global PIQA extends physical commonsense reasoning evaluation to 100+ languages and cultural contexts, revealing how language models handle everyday practical scenarios across diverse linguistic communities.

By
–
4. SmolLM2 SmolLM2 demonstrates that strategic data curation beats scale through a 1.7B parameter model trained on 11 trillion tokens using iterative data mixing optimization.

By
–
3. Multi-Agent Evolve Multi-Agent Evolve (MAE) enables LLMs to self-improve their reasoning capabilities without human-annotated data through a co-evolving multi-agent framework.

By
–
2. Introspective Awareness Anthropic research demonstrates that contemporary LLMs possess limited but functional introspective capabilities, the ability to recognize and accurately report on their own internal states.
By
–
Why CEOs Should Incentivise Employees to Replace Themselves With AI A provocative idea: it isn’t just about AI automating roles—but about leaders empowering employees to creatively use AI and transform their work, rather than being replaced by it. Read more
By
–
Grâce à l’IA, il est probable qu’en 2050 – Le cancer ne tuera plus – Alzheimer aura disparu et les mouroirs pour vieux seront un mauvais souvenir – L’université aura disparu – La sélection embryonnaire sera généralisée – Les implants cérébraux augmenteront notre
By
–
Announcing Browser Use With Scheduled Tasks
— Abacus.AI (@abacusai) 1 novembre 2025
Abacus AI's Deep Agent lets you automate all the heavy lifting associated with repetitive browser user tasks
– test your apps
– send messages on LinkedIn
– apply to jobs on a scheduled basis
Let AI work while you sleep! pic.twitter.com/HESyVEJeCz
Announcing Browser Use With Scheduled Tasks Abacus AI's Deep Agent lets you automate all the heavy lifting associated with repetitive browser user tasks – test your apps
– send messages on LinkedIn
– apply to jobs on a scheduled basis Let AI work while you sleep!