still more science showing that generative ai models are harmfully sycophantic
LLMS
-

Claude Code Integration for Automated Code Review Management
By
–
Code Review ships with: → repo-level controls
→ monthly spend caps
→ an analytics dashboard so engineering and security teams stay in charge. … and Claude never stamps a PR! Approval stays human. -
Debate: Will AI eliminate need for human code review?
By
–
This is so true.
— Linus ✦ Ekenstam (@LinusEkenstam) 9 mars 2026
Also devs thinking that vibe-coded or agentic coded solutions will need more humans to oversee and correct are 100% missing the plot.
the machines are making less and less sloppy code, and we can expect the code to become so good, humans not needed. https://t.co/bKlbyXuqbaThis is so true. Also devs thinking that vibe-coded or agentic coded solutions will need more humans to oversee and correct are 100% missing the plot. the machines are making less and less sloppy code, and we can expect the code to become so good, humans not needed.
-
Anthropic Introduces Automated AI Agent Code Review for Claude Code
By
–
🚨 Anthropic just dropped Code Review for Claude Code, and it just flipped the script on code reviews.
— Charly Wargnier (@DataChaz) 9 mars 2026
AI agents hunting bugs in parallel?
Yes please!
When a PR opens, Claude doesn't just scan the code.
It sends out multiple agents to find, cross-verify, and rank severities… pic.twitter.com/Yk2x4pCGS8Anthropic just dropped Code Review for Claude Code, and it just flipped the script on code reviews. AI agents hunting bugs in parallel? Yes please! When a PR opens, Claude doesn't just scan the code. It sends out multiple agents to find, cross-verify, and rank severities
-
Test Time Compute and Subagents: Optimizing AI Results
By
–
Roughly, the more tokens you throw at a coding problem, the better the result is. We call this test time compute. One way to make the result even better is to use separate context windows. This is what makes subagents work, and also why one agent can cause bugs and another
-
Introducing GPT 5.4 for Research Paper Analysis and Contextual Queries
By
–
Introducing GPT 5.4 for understanding research papers 🚀
— alphaXiv (@askalphaxiv) 9 mars 2026
Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references pic.twitter.com/WmWJDoK5kfIntroducing GPT 5.4 for understanding research papers Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references
-
Gary Marcus Criticizes Yann LeCun’s Glorification in AI Debate
By
–
my contempt originates in the kind of behavior I discussed here: https://
open.substack.com/pub/garymarcus
/p/the-false-glorification-of-yann-lecun?utm_campaign=post-expanded-share&utm_medium=web
… -
Gary Marcus Criticizes Yann LeCun’s Glorification in AI Discourse
By
–
see eg https://
open.substack.com/pub/garymarcus
/p/the-false-glorification-of-yann-lecun?utm_campaign=post-expanded-share&utm_medium=web
…, though it needs to be updated (again) -
Optimizing Claude Code usage for research with alphaXiv data
By
–
If you're using Claude Code for research: stop making it read directly from PDFs
— alphaXiv (@askalphaxiv) 9 mars 2026
We've introduced a SKILL.md that fetches structured, AI-friendly paper overviews from alphaXiv 👀 pic.twitter.com/AYg3n0hnB2If you're using Claude Code for research: stop making it read directly from PDFs We've introduced a SKILL.md that fetches structured, AI-friendly paper overviews from alphaXiv
