TODAY'S AI NEWS: Amazon is reportedly working on a "hybrid reasoning" model to take on AI bigwigs Plus, more news from Cohere, OpenAI, ASLP Labs, Cortical Labs, and Coreweave Here's everything you need to know:
LLMS
-
GPT-4.5 Plus tier usage limits announcement
By
–
What will the usage limits be for GPT-4.5 on the Plus tier?
-
OpenAI Rolls Out GPT-4.5 to Plus Tier Users
By
–
we are likely going to roll out GPT-4.5 to the plus tier over a few days. there is no perfect way to do this; we wanted to do it for everyone tomorrow, but it would have meant we had to launch with a very low rate limit. we think people are gonna use this a lot and love it.
-
Staggered Access Strategy for Large Language Model Conversations
By
–
so we think it's better to let people have real, long conversations with it, but that means we have to stagger people in rather than have everyone hit it hard a the same time. hope that makes sense and look forward to seeing your feedback!
-

OpenAI’s Reasoning Approach: Response Generation as Better Term
By
–
It's an OpenAI thing (
https://
openai.com/index/learning
-to-reason-with-llms/
…). I think "Response generation" would probably be a better, more general term. -

Amazon developing Nova reasoning model
By
–
Amazon is developing a reasoning model called Nova, which is expected to launch around June.
-
New Thinking Feature with Gemini 2.0
By
–
The 'Thinking' feature will require a model from the Gemini 2.0 family. It is still unclear whether it will be Flash or Pro. Full details
-
Blended Labs Uses Llama Models for Personalized AI-Native Learning
By
–
Blended Labs, an EdTech company in Germany, is using Llama 3.1 + 3.2 models to enable a wide range of AI-native flows for personalized learning pathways, real-time feedback, instant educational content generation and social gamification
-
Diffusion Forcing Combines Next-Token Prediction Video Diffusion
By
–
2. Next-token prediction meets video diffusion ▶️
— MIT CSAIL (@MIT_CSAIL) 4 mars 2025
"Diffusion Forcing" method combines the strengths of next-token prediction & video diffusion, training neural networks to handle corrupted data while predicting the next steps. This flexible, reliable sequence model helps produce… pic.twitter.com/QfsYAq7pCQ2. Next-token prediction meets video diffusion "Diffusion Forcing" method combines the strengths of next-token prediction & video diffusion, training neural networks to handle corrupted data while predicting the next steps. This flexible, reliable sequence model helps produce
-

High-Speed RAG Pipeline for PubMed Using SambaNova RDUs
By
–
Check out this high-speed RAG pipeline on @NIH #PubMed @akshay_pachaar built . He used only SambaNova RDUs, @llama_index
, and Qdrant —no GPUs! #DeepSeek #RDU