3/ A Practical Guide to Building Agents – by @OpenAI A comprehensive, step-by-step guide from OpenAI on creating and deploying AI agents. → https://
cdn.openai.com/business-guide
s-and-resources/a-practical-guide-to-building-agents.pdf
…
LLMS
-

OpenAI Releases Comprehensive Guide to Building AI Agents
By
–
-

Experiments Used GPT-4 and GPT-4o; Newer Models Could Improve Results
By
–
And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools.
-
Technical challenges in stateful AI agents and memory management
By
–
yeah they do cool stuff around memory. starting to think stateful goes beyond that to tracking changes in the agents capability, etc which I’m not sure if many memory tools do
-

LLM traces masquent un planificateur myope
By
–
“Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning” Reasoning models can write traces that look like real tree search, but this paper shows their decisions are mostly driven by shallow one-step evaluation. They extract search trees from LLM CoT in
-
Architectural challenges in building AI agent loops
By
–
agree. kinda feels like we’re tacking on memory, tools, traces, logs into agent loops but there should be an elegant way to combine them, also relying on crons or heartbeat loops also feels hack
-
Complexity of state tracking in AI agents
By
–
one thing that makes agents more complex is that they generally add capability over time which should also be tracked as part of the state
-

Critical Perspective on GPT Development and AI Capabilities
By
–

What I am about to describe ain’t AGI; it’s a sign of a trillion dollar trainwreck. If I had told you in 2022 that the 2026 version of GPT (which by the way would only be GPT 5.5 and not GPT-6 or 7 like many people fantasized about) would still have strange quirks like inserting
-
User experience comparing Claude and ChatGPT workflows
By
–
I’m increasingly using Claude (prob 50/50 now) but ChatGPT still knows me better as I’ve been using it for longer. I did try the move over info prompt which helped but not quite there yet.
-
Integrating advanced AI models into autonomous drones
By
–
Just wait for the autonomous drones to be able to use embedded Mythos-level models.
-
Debating AI model intervention in delusional thinking detection
By
–
And I think there is fertile ground for discussion inside and outside labs on ~"OK under what circumstances should we want our models to detect someone engaged in delusional thinking and intervene, and should we have same bar if they legit work at WHO or White House, etc."