Langsmith Engine is the "Full Self-Driving" moment for AI Engineering
@swyx
-
Generational upgrade of evals/analytics into continual learning platforms
By
–
every evals/analytics startup is going through a onetime generational upgrade into a continual learning platform in 2026 many will fail but as always the tasteful ones win
-

Claude Mid-Task Instruction Updates Without Cache Break
By
–

"Developers can update Claude’s instructions mid-task without breaking the prompt cache or routing the update through a user turn" wtf? how??
-

Devin Smart Automations: First Non-Annoying Proactive Agent
By
–
oh yeah first* automations – instead of dumb crons kicking of agents, devin has made them smart. try it – it is the first non annoying proactive agent impl ive seen
-

Cognition: Largest Independent Agent Lab Growth Analysis
By
–

cognition is now the largest independent agent lab in the world. take the 200% utilization that everyone is hitting from this chart and run out the sales growth from this, i encourage you to go thru the exercise if you are new to investing a lot of you have read my cog
-

Optimizing Codex Prompts for AI Agent Automation
By
–
UPDATE: Came up with an even better version of this prompt after the feedback Ask Codex to look across your sessions, Memories, and Chronicle, identify patterns, reuse what already exists, and only create the smallest useful skill, subagent, or automation. "Look back over my x.com/reach_vb/statu…
-

Transformer Learning Frameworks and Adversarial World Models
By
–
co-sign. a very handy mental framework for what kinds of learning transformers do well today, and why it runs into limitations. when @ankit2119 and i wrote about the need for adversarial world models earlier this year, we were describing a couple of the functions of these rungs
-
AI Agents for Automated Code Hardening and Audit
By
–
Kakuna: skills with checklists that only know how to harden your codebase /plan with it then let it /goal for a day, it comes back with same functionality but all the boring stuff done for you + an audit of its own work. focus on subagent parallelism and encodes strong
-

Vibecoded App to Production-Ready AI Agent Repo
By
–
working on a "take this vibecoded slop app and make it a production-ready, e2e tested, maintainable, parallelizable agent repo" skill. this thing ran for ~16 hours yesterday and made 103 commits all told and i ended up with exactly the same app but instead of fragile mvp it
-
Correlation Between Model Performance and AI Agent Revenue
By
–
very belated but in retrospect i think @sama's mythical "build a business that gets better when models get better" is basically what I called Agent Labs here.
— swyx (@swyx) 20 mai 2026
seeing a very direct correlation with model performance and agent lab revenue, discontinuity in Q4 2025
(clip from… https://t.co/ffp1RBLlhL pic.twitter.com/YF8Xuml8hhvery belated but in retrospect i think @sama
's mythical "build a business that gets better when models get better" is basically what I called Agent Labs here. seeing a very direct correlation with model performance and agent lab revenue, discontinuity in Q4 2025 (clip from