@baseten users are scaling smarter with us: 5× throughput on high-traffic endpoints 50% lower cost per token Up to 38% lower latency on the largest LLMs Built on NVIDIA Blackwell + TensorRT-LLM + Dynamo on @googlecloud
—driving efficiency, speed & adoption at scale.
LLMS
-

Baseten Achieves 5x Throughput Scaling on LLM Endpoints
By
–
-
LangGraph: Essential Features for Production AI Agents
By
–
Building AI agents for production comes with unique challenges. In our latest blog, we share how we designed LangGraph to tackle them: Why heavy abstractions fail and what really matters for control & durability The 6 features every production agent needs in practice
-

Frontier Models 90% Cheaper Yet 2.7x Better Performance
By
–
Not sure how many people would have taken the bet 2 years ago that by now frontier models would be over 90% cheaper but 2.7x better. Today it's simple fact. The tech gains are cool, but it's the leap forward in access that makes it meaningful.
(graph: @ethanmollick) -
Groq Launches Compound Agentic System for Production at Scale
By
–
📣Groq’s first agentic system is ready for production at scale. Already battle tested by 100K+ developers across 5M+ requests.
— Groq Inc (@GroqInc) 4 septembre 2025
Compound is now GA, available to everyone on GroqCloud.
Go Build ⬇️ pic.twitter.com/vbBCavHySnGroq’s first agentic system is ready for production at scale. Already battle tested by 100K+ developers across 5M+ requests. Compound is now GA, available to everyone on GroqCloud. Go Build
-
LLMs and the impact on entry-level job tasks
By
–
There could be other effects! For me it's quite intuitive that entry-level jobs are more affected, because it's the jobs whose tasks are generally lower-level, so more accessible to LLMs
-
Qwen to Release More Intelligent Model, China’s AI Progress
By
–
It seems as if Qwen is going to release another, even more intelligent model. China is on fire!
-
kimi-researcher: game changer, unbeatable free tool
By
–
kimi-researcher is a game changer, can't beat free
-
ChatGPT sabotaged by OpenAI: complete analysis of the quality drop
By
–
#ChatGPT has become worse… And it's done on purpose! Why did @OpenAI sabotage its best AI? I explain everything in this complete analysis: https://youtu.be/jk0UDNmH8zs #GPT5 #ChatGPT5 #ChatGPTdown
-

Mistral AI Raises $2.3 Billion at $14 Billion Valuation
By
–
ACTU : Mistral AI est en train de finaliser un investissement de 2,3 milliards de dollars pour une valorisation de 14 milliards.
-
MiniCPM-V 4.5 Achieves 77.0 Score Surpassing GPT-4o and Gemini
By
–
MiniCPM-V 4.5
— AK (@_akhaliq) 4 septembre 2025
achieves an average score of 77.0 on OpenCompass, a comprehensive evaluation of 8 popular benchmarks. With only 8B parameters, it surpasses widely used proprietary models like GPT-4o-latest, Gemini-2.0 Pro, and strong open-source models like Qwen2.5-VL 72B
powered… pic.twitter.com/H7s4zanvSKMiniCPM-V 4.5 achieves an average score of 77.0 on OpenCompass, a comprehensive evaluation of 8 popular benchmarks. With only 8B parameters, it surpasses widely used proprietary models like GPT-4o-latest, Gemini-2.0 Pro, and strong open-source models like Qwen2.5-VL 72B powered