LLM Agents in Production: Architectures, Challenges, and Best Practices – ZenML Blog https://
bit.ly/4gGcPtW
#AI #MachineLearning #DeepLearning #LLMs #DataScience
LLMS
-

LLM Agents Production: Architectures, Challenges, Best Practices
By
–
-

Fireside Chat on LLM-Based Customer Support Agent Systems
By
–
If you love diving deep into the technical details of highly complex, continually improving LLM-based systems… then boy do I have an event for you! Next Tuesday, in SF, fireside chat with Decagon – one of the leading customer support agent builders https://
lu.ma/w7y0bqwr -

DeepSeek-R1 Open Source Model Available as NVIDIA NIM Microservice
By
–
The open source DeepSeek-R1 model is now available as an NVIDIA NIM microservice preview on http://
build.nvidia.com to help developers securely experiment with its advanced AI reasoning capabilities. -
Deepseek models now available in Cursor editor
By
–
Deepseek models are available now in Cursor! Hosted on US servers. While we're big fans of Deepseek, Sonnet still appears to perform much better on real-world tasks. Enjoy!
-
Enable DeepSeek R1 and V3 Models in Settings
By
–
To enable, head to Settings > Models. Then, check deepseek-r1 or deepseek-v3.
-
Dario’s DeepSeek Essay: Closed-Source Justification Critique
By
–
Finally took time to go over Dario's essay on DeepSeek and export control and to be honest it was quite painful to read. And I say this as a great admirer of Anthropic and big user of Claude* The first half of the essay reads like a lengthy attempt to justify that closed-source
-
Google and Meta dominate AI model pre-training landscape
By
–
Thanks! Yeah seems like so far Google/Meta are the only two adtech horses in the race for pre-training. (I'm curious if any ML engineers with Llama/Gemini experience also have ad-tech experience.) Would love to hear more what you're doing with post-training though!
-

Alibaba’s Qwen 2.5 Max surpasses DeepSeekR1, free video test
By
–
#DeepSeekR1 drives the world crazy but no, it's not the best Chinese #ChatGPT! @AlibabaGroup just released its AI, Qwen 2.5 Max, and it's just crazy. I test it in my latest video right here → https://youtube.com/watch?v=zE_dM7DwPQw … Everything is free… Even generate videos.
-
LLM Training Experience in AdTech Products
By
–
Anyone in #adtech (currently or previously) have experience training LLMs, either open-source or closed models? Also looking to connect with more people that have direct experience building adtech products using LLMs.
-
AI Paradigm Shift: From Imitation Learning to Reward Learning
By
–
This is a monumental shift in AI. We’re moving from a world of imitation learning (with SFT), to reward learning (with RL). AI that not just creatively imitates images and language but truly invents new strategies and insights.