Yes, next to residual attention as hot new candidate. And, I am also of course curious to see what DeepSeek V4 is up to haha
RESEARCH
-

Machine Learning with Python for Everyone Book Release
By
–
Machine Learning with Python for Everyone: http://
amzn.to/43HEaFS
(part of the Data & Analytics Series from Addison–Wesley Publishers)
—————
#ML #DataScience #AI #DataScientist -
Cross Attention in Separate Articles on Multi-Modal Models
By
–
Yeah, I have cross attention in my separate articles (the multi-modal and understanding attention from-scratch) ones. This article was focused on text LLMs to keep the scope reasonable.
-
XSA Attention Mechanism: Future Research Candidate for Follow-up Study
By
–
Nice one! XSA from https://
arxiv.org/abs/2603.09078? To keep the scope reasonable, I only focused on those that already made it into the flagship architectures. But maybe that's an interesting candidate for a follow-up on attention research candidates. -

ARO Framework Dramatically Accelerates LLM Training via Gradient Rotation
By
–
Can LLM training be dramatically accelerated beyond current methods? Microsoft Research, The Chinese University of Hong Kong, Shenzhen, and University of Wisconsin-Madison introduce ARO. This new matrix optimization framework pioneers "gradient rotation" as a core principle.
-

Using Hugging Face Papers API as an AI Development Skill
By
–
If you're building anything in AI, the best skill you need to be using right now is hugging-face-paper-pages Whatever problem you're facing, someone has probably already published a paper about it. HF's Papers API gives a hybrid semantic search over AI papers. I wrote an internal skill, context-research, that orchestrates the HF Papers API into a research pipeline. It runs five parallel searches with keyword variants, triages by relevance and recency, fetches full paper content as markdown, then reads the actual methodology and results sections. The skill also chains into a deep research API that crawls the broader web to complement the academic findings. The gap between "a paper was published" and "a practitioner applies the insight" is shrinking, and I think this is a practical way to provide relevant context to coding agents. So you should write a skill on top of the HF Paper skill that teaches the model how to think about research, not just what to search for.
→ View original post on X — @thom_wolf, 2026-03-22 18:35 UTC
-
Balloon-Powered Robots: Innovation in Robotics and Engineering
By
–
Balloon-powered #Robots
— Ronald van Loon (@Ronald_vanLoon) 22 mars 2026
by @DennisHongRobot
#Robotics #Engineering #ArtificialIntelligence #Innovation #Technology pic.twitter.com/GMuGM2hExZBalloon-powered #Robots
by @DennisHongRobot #Robotics #Engineering #ArtificialIntelligence #Innovation #Technology -
Immigration Policy Critical for US AI Competitiveness Against China
By
–
If you really want the US to have an AI advantage over China, immigration policy is at least as important as chip controls. Attract the best and brightest. Keep them here. Don't make it easier for them to do AI work in China than the US. Talent is vital. Zephyr (@zephyr_z9) 40%-60% of the top researchers at frontier labs are non-US citizens If they get removed, then American labs will lose the race — https://nitter.net/zephyr_z9/status/2035759774017667291#m
-

Interactive Introduction to Quadtrees for AI Applications
By
–
An interactive intro to quadtrees https://
buff.ly/5m2COK8
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Pedro’s Theoretical Insights on Intelligence Beyond Reinforcement Learning
By
–
This is another outstanding theoretical step towards understanding intelligence by Pedro, aka @AdaptiveAgents Some people still like to try to explain it all through RL, but I feel other explanations that emphasise the role of the environment, entropy minimisation, (multi)