Impressive benchmarks!: patching together agents really improves perf on many benchmarks, à la Manus or OpenHands-Versa
AGENTS
-
OpenAI Presents ChatGPT Agent with Action Execution
By
–
🔴 ¡OPENAI presenta CHATGPT AGENT!
— Carlos Santana (@DotCSV) 17 juillet 2025
La nueva capacidad de ChatGPT de navegar y ejecutar acciones en internet, y razonar durante más tiempo usando las herramientas adecuadas (generación y ejecución de código, generación imágenes, etc.) para cumplir su tarea
Os voy contando! 👇🧵 pic.twitter.com/xldxVmf8uO¡OPENAI presenta CHATGPT AGENT! La nueva capacidad de ChatGPT de navegar y ejecutar acciones en internet, y razonar durante más tiempo usando las herramientas adecuadas (generación y ejecución de código, generación imágenes, etc.) para cumplir su tarea Os voy contando!
-
Example of web research, Drive, Terminal slides, and Operator validation
By
–
Example of cool capabilities unlocked by this : do research on the web and connect to Drive via Text browser, then create slides via Terminal, then view them through Operator to validate slides or go back to the drawing board.
-
Deep research and Operator limitations; Operator could do deep research
By
–
The team acknowledges that some agents are still very limited each in their own way:
– Deep research cannot interact with webpages (no Textbrowser can)
– Operator has trouble reading through long pages personal take: Operator should theoretically be able to also do deep research -

ChatGPT agent live announcement combining Deep Research, Operator, Terminal code editing
By
–
ChatGPT agent live announcement underway, with @sama onboard!
This Agent patches together
– Deep Research (TextBrowser )
– Operator (GUI Agent) – Terminal code editing
to unlock all these capacities together in an agent. -
Manus vs o3: Comparing AI Agent Capabilities
By
–
It feels like a cross between Manus and o3. Manus is capable of somewhat more complex tasks, but agents is better at integrating research and doing a wider range of tasks well (this is early impressions, and I know they are still working on the tool)
-
AI Paradigm Shift: From Prompting to Task Delegation
By
–
It feels much more like working with an actual human intern capable of a wider range of analytical and computer tasks, and, like an intern, you want to give it feedback and work back & forth. Not all the way there yet, but the paradigm is shifting from prompting to delegating.
-

ChatGPT Agents Enable Autonomous Research and Document Creation
By
–
I had early access & ChatGPT agent is, I think, a big step forward for getting AIs to do real work Even at this stage, it does a good job autonomously doing research & assembling Excel files (with formulas!), PowerPoint, etc. It gives a sense of how agents are coming together
-

OpenAI prepares a unified agentic model
By
–

OpenAI is about to announce a “Unified Agentic Model” Will it be GPT-5 in fact? The one to be released at September? Wondering if “Agentic” may mark another tier after “Reasoning” models.
