opens X and sees @OpenAI just dropped o1 in Canvas on Friday afternoon… I AM CALLING FOR A COMPLETE AND TOTAL PAUSE ON AI PROGRESS UNTIL MONDAY. I AM TIRED
LLMS
-
Using R1 Distills Models Without RAG Implementation
By
–
no rag, but you can use r1 distills (the smaller ones)
-
Windows 95 Simulator Built in ChatGPT with Canvas and o1
By
–
One of the coolest examples I’ve seen: @edwinarbus instantly built a Windows 95 simulator in ChatGPT with canvas using @OpenAI o1—then started playing Minesweeper right there! It just worked! https://
x.com/edwinarbus/sta
/edwinarbus/status/1882877195930202384
… -
Best ArXiv Papers Read Throughout 2024
By
–
Ha, yeah, it's basically the "best of" of the arxiv slice I read in ~365 days in 2024 😛
-
ChatGPT o1 Now Renders Interactive React Apps in Canvas
By
–
ChatGPT now lets you build and render interactive front-end experiences with @OpenAI o1—right in the canvas! No more just outputting code, you can create and run your React apps all in one place! https://t.co/Lv7U2LnO8m
— Romain Huet (@romainhuet) 24 janvier 2025ChatGPT now lets you build and render interactive front-end experiences with @OpenAI o1—right in the canvas! No more just outputting code, you can create and run your React apps all in one place!
-
PhD Research on Attention Architecture and NLP Benchmarks
By
–
This gave me a good chance to reflect on my phd years, it was a pretty fun time. I mostly did tons of architecture work, played around a lot on attention blocks, recurrence etc. Felt like a professional lego builder. I was playing on the SNLI, MNLI, Squad leaderboards etc,
-

PhD Thesis Critique Using Gemini 2.0 Flash Thinking
By
–
Trying out this fun critique trend on my PhD thesis ("neural architectures for natural language understanding",) using gemini 2.0 flash thinking. It thinks most of my thesis is empirical and has no theoretical contribution (yes, true), no strong overarching narrative cohesion
-
Llama Stack First Stable Release with API Improvements
By
–
Today, we’re publishing the first stable release of Llama Stack.
With this release Llama Stack now includes:
• Streamlined upgrades w/ backwards compatibility for future API versions.
• Automated verification for supported providers. Live in the repo -
Apollo Introduces Thinking Tokens Latest Version
By
–
Thinking tokens out now in the latest version of Apollo.
-
DeepSeek R1 1.5B Achieves 4o Performance on Any Hardware
By
–
Here’s Deepseek r1 1.5B thinking through a problem — it’s comparable to 4o and Claude 3.5 Sonnet in a number of domains like math. Except…
— Aaron Ng (@localghost) 24 janvier 2025
it’s a 1.5B model…
and can run on virtually any hardware. Truly a huge efficiency leap. pic.twitter.com/CjvsCaGiU3Here’s Deepseek r1 1.5B thinking through a problem — it’s comparable to 4o and Claude 3.5 Sonnet in a number of domains like math. Except… it’s a 1.5B model… and can run on virtually any hardware. Truly a huge efficiency leap.