Me chatting with Karen Simonyan on my first day at @Microsoft AI: “… no way! That’s such a brilliant idea… so simple … wow, I can’t believe no one saw it before”. This is what I love the most about my field. Just when you think it’s gonna take gargantuan effort or feel
@nandodf
-
Microsoft AI Hiring Exceptional Engineers for Multimodal Research
By
–
I’ve joined @Microsoft AI to advance the frontier of large scale multimodal AI research and to build products for people to achieve meaningful goals and dreams. The MAI team is small, but well resourced and ambitious. We are now looking for exceptional ICs, who like to ship. If
-
Deep Indaba and Khipu AI: Models for AI Education
By
–
I very much look forward to continuing to participate in the @DeepIndaba
. Helping you, Shakir and Ulrich start it was a very enriching experience. It gave me so much, and it keeps giving. The deep learning Indaba became a model for @Khipu_AI and many other amazing education -
California AI Policy Decisions and Geopolitical Impact Analysis
By
–
I would love to hear the opinion of game theorists and geopolitical experts on this. I worry that this California choice will impact all of us, and that California, although amazing, is not always the best at solving problems (eg dire state of homelessness in SF). I find it hard
-
Researcher Departs Google DeepMind After Decade of AI Advancement
By
–
It’s time to say thank you and goodbye to @GoogleDeepMind
. I had the immense fortune of working there for 10 years. They were undoubtedly the most exciting years in the history of AI, and I feel that I grew beyond all my expectations thanks to my uniquely smart, generous and -

Llama 3 Paper: Essential Reading for Building Leading LLMs
By
–
The Llama 3 paper is a must-read for anyone in AI and CS. It’s an absolutely accurate and authoritative take on what it takes to build a leading LLM, the tech behind ChatGPT, Gemini, Copilot, and others. The AI part might seem small in comparison to the gargantuan work on *data*
-
Team Leadership Over Individual Voice in AI Research
By
–
Chef Wisdom for AI Researchers 2: ‘I believe in the team, the team is what I protect, not the individual. In fact, I think that they are giving a little too much voice to chefs. Shut up and work…It is a formula that has always worked very well for me.’ Albert Adrià [maybe
-
Growth Through Uncertainty: Innovation at the Edge
By
–
Chef Wisdom 1: "To grow, to improve, you have to be there at the edge of uncertainty." -Francis Mallmann
-
Matrix Abstraction Trade-offs: Balancing Speed and Context Tracking
By
–
Fair, Sasha. For me, it’s like playing with matrices. The matrix abstraction is clearly useful for fast manipulation. However, sometimes I find it helpful to keep detailed track of indices and context. I agree that sometimes too much context can make something confusing, and
-

Training 124M Parameter LLMs on MacBook M3 in Real-Time
By
–
It is remarkable that anyone can now train a 124M parameter LLM in about real-time on a MacBook M3. So easy to experiment. This would have been the stuff of dreams when I was in school. I training neural nets, but I really admire the people who build the hardware.