There's a difference between using AI and building with it. Copy-pasting ChatGPT prompts will only get you so far. I want to help you learn to build personal AI software, automations, and tools that actually solve your problems. The AI Fast Track is a free 5-day course. Tens of
CODE
-
FFmpeg’s Firm Stance on Open Source Development Funding
By
–
I love how @FFmpeg is conducting themselves online. It's a project that's created billions of dollars in actual value and has captured practically none of it. Their stance seems to be simple: fund substantial development or STFU with the requests.
Pretty easy to understand. If -

MLflow-Powered Domain-Specific Judges for Agent Bricks
By
–
Building accurate, trustworthy agents in Agent Bricks starts with better judges — because AI agents can only be as good as the ones evaluating them.
— Databricks (@databricks) 4 novembre 2025
We’re excited to streamline creation of domain-specific judges with new MLflow-powered capabilities in Agent Bricks:
•… pic.twitter.com/vEDE2Rabg6Building accurate, trustworthy agents in Agent Bricks starts with better judges — because AI agents can only be as good as the ones evaluating them. We’re excited to streamline creation of domain-specific judges with new MLflow-powered capabilities in Agent Bricks:
• -

IMO Medalists Grade AI Homework with Human Verification
By
–
We do have teachers (IMO medalists) to grade the homeworks too 🙂 See the paper https://
arxiv.org/abs/2511.01846 where we recommend to augment with human verifications. -

MiniMax-M2 Open Source Model Available on Poe
By
–
MiniMax-M2 is now available on Poe! This open source model has a 200k token context window, has 230b parameters with an MoE architecture, and excels at coding and agent workflows. (1/2)
-

ProofAutoGrader: Automatic IMO Proof Evaluation Using Gemini
By
–
While human expert evaluation remains the gold standard for mathematical proofs, its cost and time intensity limit scalable research. To address this, we built #ProofAutoGrader, an automatic grader for IMO-ProofBench. The autograder leverages Gemini 2.5 Pro, providing it with a
-
LangChain Releases Human-in-the-Loop Agent Middleware Series
By
–
Over the next few weeks we'll be releasing a series of deep dive videos into our new prebuilt agent middlewares. We're starting off with one the most popular middlewares, human-in-the-loop! Require approval on sensitive tool calls before execution in just 1 LOC!
-
Waterfall Development with LLMs: Inherent Challenges
By
–
Waterfall based development with LLMs have all the challenges with waterfall based dev?
-
LangSmith for Improving Agent Quality at Scale
By
–
Why we built LangSmith for improving agent quality
— LangChain (@LangChain) 4 novembre 2025
As more agents move into production, teams need to move beyond vibe-checking and bring rigor to how they understand agent behavior at scale.
In this video, the LangSmith engineering team and Harrison (@hwchase17) sit down to… pic.twitter.com/q7jrXfWJo1Why we built LangSmith for improving agent quality As more agents move into production, teams need to move beyond vibe-checking and bring rigor to how they understand agent behavior at scale. In this video, the LangSmith engineering team and Harrison (
@hwchase17
) sit down to