Agent skills are now available in the Deep Agents CLI, enabling you to use the large and growing collection of public skills with your agents. In this video we discuss: – What agent skills are and why they’re interesting
– How agents make use of skills
– How you can use skills
AGENTS
-
Agent Skills Now Available in Deep Agents CLI
By
–
-

Building Multi-Agent Systems with Vision Capabilities
By
–
LLMs can't see. How can we build effective multi-agent systems with vision capabilities? Building multimodal models from scratch is expensive. Training joint vision-language architectures requires massive compute, specialized datasets, and careful optimization. But there's
-
VLM Benchmarking for Long Horizon Robotic Household Tasks
By
–
Our most recent work that benchmarks modern VLM and their efficacy for long horizon household activities in robotic learning, using BEHAVIOR benchmark environment.👇 https://t.co/8Ibx8IA7MW
— Fei-Fei Li (@drfeifei) 25 novembre 2025Our most recent work that benchmarks modern VLM and their efficacy for long horizon household activities in robotic learning, using BEHAVIOR benchmark environment.
-

AI21Labs and DeepChecks Host AI Agents Meetup at AWS reInvent
By
–
Heading to @awscloud #reInvent? Join @AI21Labs & @deepchecks for an AI agents meetup on Dec 4 in Las Vegas. Learn from real GenAI deployments and meet teams tackling similar challenges. Spots are limited, register here: https://
ai21.com/events/reinven
t-2025/?utm_source=org-twitter
… -

Generative AI Tech Stack Framework for Autonomous Agents
By
–
The Generative AI ecosystem is evolving into a full tech stack — powering autonomous AI agents.
From infrastructure and LLMs to RAG pipelines, agent behaviors and orchestration layers, this framework shows the 6 layers driving next-gen AI systems. Credit: @goyalshalini #AI -
Claude’s Task-Level Savings Estimates: Current Limitations and Future Improvements
By
–
Our study has limitations: above all, Claude can’t use what happens outside of the chat window to refine its estimate of task-level savings. But as models improve, we think its estimates of task-level savings will improve too. We’ll return to this research soon.
-

Claude’s Task Duration Estimates Show Promise Over Time
By
–
We first tested whether Claude can give an accurate estimate of how long a task takes. Its estimates were promising—even if they’re not as accurate as those from humans just yet.
-
AI Plays StarCraft with Robot Mouse and Keyboard Control
By
–
Happy to play against it on @StarCraft (former pro). Bonus points if you get the robots to operate the mouse and keyboard
-
Five AI Agent Use Cases Transforming Business in 2026
By
–
5 Amazing AI Agent Use Cases That Will Transform Any Business In 2026 #AIagents represent the evolution beyond simple #chatbots, capable of autonomously planning and executing complex business workflows from start to finish. This article explores five practical applications,
-
Abacus AI Launches High-Effort DeepAgent with Top Models
By
–
Opus 4.5 is so good that we will be launching a “high effort” version of the Abacus AI DeepAgent tomorrow! It will incorporate Opus 4.5 alongside Gemini 3 and GPT 5.1 All the top models working together to get your job done!