Stress-testing model specifications, led by Jifan Zhang. Generating thousands of scenarios that cause models to make difficult trade-offs helps to reveal their underlying preferences, and can help researchers iterate on model specifications.
AGI
-

Anthropic Fellows Program Releases Four AI Safety Research Papers
By
–
The Anthropic Fellows program provides funding and mentorship for a small cohort of AI safety researchers. Here are four exciting papers that our Fellows have recently released.
-
AGI Definition: Feature-Complete Remote Worker with High Reliability
By
–
When I say AGI I mean a "feature complete" build of a remote worker with a lot of 9s (think: a bit like current Autopilot FSD, maybe plus a few more iterations). This is the original definition I've stuck to forever. Separate from the diffusion /implementation of it across
-
The Evolution of AI from Tools to Thinking Collaborators
By
–
Denario shows what happens when reasoning, experimentation, and authorship merge into one system. It doesn’t just help humans do science faster it expands what science can be. We’re not building tools anymore.
We’re building thinking collaborators. Are you excited or scared? -
What capabilities does AI still need to develop?
By
–
What’s one thing you’d like AI to be able to do, but it can’t yet?
-

GAP: Graph-Based Agent Planning with Parallel Tool Execution
By
–
6. GAP GAP introduces graph-based agent planning with parallel tool execution and reinforcement learning, enabling AI agents to coordinate multiple specialized capabilities simultaneously rather than sequentially.
-
Meta and the collaborative pursuit of AGI
By
–
Meta is doing the hard work here. I appreciate it. To achieve general intelligence… all these company must get together.
-
French political programs ignore incoming superintelligence threat
By
–
La totalité des programmes politiques en France ignorent l’arrivée prochaine de la Super Intelligence Artificielle Derrière le cirque à l’Assemblée, nous préparons un déclin accéléré de la France Nos enfants et nos petits-enfants nous détesteront
-
Recursive Language Models for Near-Infinite Context Agents
By
–
One of our best talks yet. Thanks @a1zhang for the amazing presentation + Q&A on Recursive Language Models! If you're interested in how we can get agents to handle near-infinite contexts, this one is a must. Watch the recording here!
-
Agents vs Workflows: Breaking Down AI Architecture Patterns
By
–
You could technically make an agent into a workflow by having the workflow have LLM steps which dynamically route… but it starts to get all confusing that way, and the user experience is also more tricky. I've started to like the way @AnthropicAI breaks it down: * Skills
