MCP is awesome. Kudos to @dsp_ and team for creating this standard.
AGENTS
-
Encrypted Thoughts in Multi-Turn Tool Calling Conversations
By
–
For tool calling, we are working on a way to have raw encrypted thoughts in multi-turn conversations, working with different teams right now to make sure this is battle tested!
-
The impact of human-like AI agents on customer service and media
By
–
AI agents that sound human and can handle emotion, pause, interruption…. This is going to hit customer service, gaming, and podcasting hard. Big shift ahead.
-
BrowseComp Benchmark for Evaluating Agent Search Capabilities
By
–
Interesting launch! If the agent is good at "agentic search that doesn't stop until it finds what you need", consider evaluating on our BrowseComp benchmark, which measures just that! SimpleQA mainly targets models that don't browse: https://
openai.com/index/browseco
mp/
… -

New Agent Robustness and Control Team Recruitment
By
–
join our new Agent Robustness and Control team:
-
Claude Projects Feature Rolls Out to Paid Plans
By
–
Rolling out to all paid Claude plans over the coming days. Try it at https://
claude.ai/projects. -
Why AI agents fail at programming: scaffolding solutions
By
–
My new dev diary is out: Why AI agents fail at programming (hint: wrong scaffolding) Is $100/day for AI coding expensive? Not if it replaces 2 weeks of work SE2: How we're making complexity fun instead of overwhelming Read the full diary:
-

AI Personalized Learning: Transforming Education for Every Student
By
–
If you’re lucky, you’ve come across a teacher or mentor in your life who just *got* how you learn. What if AI could do that – for everyone? Research is rolling in on what a difference that could make for students. Take this paper from the CHI conference. 1/