What it actually feels like to be using AI agents to build an app. pic.twitter.com/ewx1XMYTnL
— Bojan Tunguz (@tunguz) 19 janvier 2026
What it actually feels like to be using AI agents to build an app.
By
–
What it actually feels like to be using AI agents to build an app. pic.twitter.com/ewx1XMYTnL
— Bojan Tunguz (@tunguz) 19 janvier 2026
What it actually feels like to be using AI agents to build an app.
By
–
Now you can run many SWE agents like Claude Code, Codex CLI, Gemini CLI and others on remote VMs, managed with a single Agents API released by @blackboxai
— π¨ AI News | TestingCatalog (@testingcatalog) 19 janvier 2026
AI Agents outsource π€ https://t.co/MMUN4SPlwC pic.twitter.com/I0k7v6BDri
Now you can run many SWE agents like Claude Code, Codex CLI, Gemini CLI and others on remote VMs, managed with a single Agents API released by @blackboxai AI Agents outsource
By
–
DeepReinforced announced InterX, a new AI system for Deep Code Optimisation based on Reinforcement Learning.
— π¨ AI News | TestingCatalog (@testingcatalog) 19 janvier 2026
Suitable for optimising:
– CUDA Kernels
– Smart Contract
– SQL queries
– And more! π https://t.co/PzmdOQQSQR pic.twitter.com/H4lBYURcjW
DeepReinforced announced InterX, a new AI system for Deep Code Optimisation based on Reinforcement Learning. Suitable for optimising:
– CUDA Kernels – Smart Contract
– SQL queries
– And more!

By
–
Top stories in AI today: – OpenAI officially bringing ads to ChatGPT
– The Rundown Roundtable: Our AI use cases
– Code from your phone with OpenAIβs Codex
– Musk, OpenAI trade (more) public blows
– 4 new AI tools, community workflows, and more Read more: https://
therundown.ai/p/ads-are-offi
cially-coming-to-chatgpt
β¦

By
–
What DIED in 2023-2026: Prompt Engineering: β60% salary ($95Kβ$38K). 100K+ certified, jobs disappeared. Basic Python: β61% ($82Kβ$32K). AI writes better code. Entry Data Analysis: β64% ($78Kβ$28K). Automated dashboards replaced analysts. Commoditization happened
By
–
here's the part that stuck with me. claude outperformed every human candidate who's ever taken anthropic's internal engineering hiring test. not most. every single one in the company's history. a 2-hour rigorous assessment designed to filter for elite engineers. so what

By
–
the benchmark data made me pause. claude opus 4.5 hit 80.9% on SWE-bench verified. first model to ever break 80%. this isn't leetcode. it's real github issues from production repos. the actual work developers do. 4 out of 5 real-world bugs. solved.

By
–
i looked up the internal numbers. their engineers are using claude for 60% of their work. not "sometimes helps with debugging." sixty percent of everything. 50% productivity gains. 2-3x better than a year ago.
at what point do we stop calling this "assistance"?

By
–
daniela amodei on CNBC: "by some definitions of that, we've already surpassed" human-level AI. her example was coding. she said claude writes code "about as well as many developers at anthropic now." anthropic. the company that probably employs some of the best engineers on

By
–
lol feels like the team introduced this bug just to troll users