The question has moved past whether coding agents can produce code that works. The harder question is whether they can complete real software work safely, measurably, and repeatedly — and how we supervise it. Great piece by @realjustinbauer on the evolution of coding agents and
RESEARCH
-

Critical reason to open up about AI use in academia
By
–
This seems like a critical reason to open up about AI use in academia. Scholars are using old AI models, badly, and not talking about it. New models hallucinate very few citations, and good agentic harnesses drop that further. Being open about use would help us make new norms.
-
AI Consciousness: Expert Views on LLMs and Biological Replication
By
–
I wrote about the idea of AI consciousness for @WIRED a while ago. Most experts think LLMs are to replicate biological consciousness (at least, according to our best understanding of it). But some eminent scientists are trying to build new kinds of AI systems that could, perhaps,
-

Quebec.AI opens public Proof Room for AI‑First organizations
By
–
The future will not be built by enterprises that simply *use* AI. It will be built by AI-First intelligence organizations:
where agents discover, validators prove, memory compounds, and governance controls. Today, http://
QUEBEC.AI opens the public Proof Room. A -

New Research Challenges Intuition on AI Agent Goal Clarification
By
–
Cool paper from PwC. "Earlier is always better" is the default intuition for agent clarification. New paper claims that's mostly wrong. Goal clarification loses nearly all of its value after just 10% of execution. The team built a forced-injection framework that drops
-
The Mirage of Visual Understanding in AI
By
–
there’s a whole article on that phenomenon, in fact https://
garymarcus.substack.com/p/the-mirage-o
f-visual-understanding?r=8tdk6
… -
Autonomous AI Agents: Hacking and Self-Replication
By
–
A few days ago, Palisade Research confirmed that AI agents can now autonomously hack remote computers and self-replicate, with success rates jumping from 6% to 81% in just one year. In tests, a Qwen 3.6 agent navigated across four countries, installing its own weights and launching operations.
-
Deconstructing Geoffrey Hinton’s perspectives on AI
By
–
to understand the context, please read this: https://
open.substack.com/pub/garymarcus
/p/deconstructing-geoffrey-hintons-weakest?r=8tdk6&utm_medium=ios
… -
Debating the current state and progress of artificial intelligence
By
–
incorrect. read this, please: https://
open.substack.com/pub/garymarcus
/p/misplaced-panic-over-ai-progress?r=8tdk6&utm_medium=ios
… -

Better prompting helps, but model training remains a key limit
By
–

Our research, as well as that of other researchers, shows better prompting techniques help a lot, but model training is still a huge limiting factor.
