
Deepseek-v3-0324 is now also available on @lmarena_ai for you to try!

By
–

Deepseek-v3-0324 is now also available on @lmarena_ai for you to try!

By
–
DeepSeek-V3-0324 dropped on @huggingface
, an upgraded version of a previously released V3 version.

By
–
Guest Lecture 6 CS329A by Prof. Jeff Clune:Open-ended Agent Learning in the Era of Foundation Models https://
buff.ly/ILupGfY
#AI #MachineLearning #DeepLearning #LLMs #DataScience

By
–
DeepSeek just rolled out an update for its V3 model! DeepSeek-V3-0324
Hugging Face: https://
huggingface.co/deepseek-ai/De
epSeek-V3-0324/tree/main
…

By
–
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't Quy-Anh Dang, Chris Ngo: https://
arxiv.org/abs/2503.16219 #DeepLearning #ChatGPT #ReinforcementLearning

By
–
Understanding Attention in LLMs https://
buff.ly/qg9OeRP #AI #MachineLearning #DeepLearning #LLMs #DataScience
By
–
Congrats to @eugeneyan for yet another banger ranking on HN! Join the Latent Space Paper Club this wednesday for live Q&A with him! https://
lu.ma/n6w3mijk Recsys + LLMs is dynamite!

By
–

Anthropic keeps working on its "Compas" feature and adding a new toggle to the updated composer UI. Assumingly, "Compass" will allow Claude to perform certain tasks and likely will be similar to Deep Research.
By
–
“Today’s AI models are static. Once deployed, they do not change when they are presented with new information. This is a remarkable shortcoming for any intelligent system to have. It represents a profound weakness of artificial intelligence compared to biological intelligence.”