The problem is not that science leaders deceived some people, but that none of those within the institutions who were not deceived stood up and contradicted them
ETHICS
-

Claude’s Multifaceted Capabilities Across Human and Technical Domains
By
–
“Claude, change a diaper, plan an invasion, butcher a hog, conn a ship, design a building, write a sonnet, balance accounts, build a wall, set a bone, comfort the dying, take orders, give orders, cooperate, act alone, solve equations, analyze a new problem, pitch manure, program
-
Science integrity compromised by opportunists and ideologues
By
–
Science requires the adherence to a special code, a commitment to seek and serve truth above all other things. By opening up science to opportunists, impostors and ideologues, we defiled it, robbed it of its status and sanctity, and deprived ourselves of its fruits
-
Commercial vs Personal AI Use: Ethical and Legal Boundaries
By
–
If for fun it's fine, commercially I don't allow it
-

Image blocking issues with AI content moderation systems
By
–
Trying!!! But everything is getting blocked, tried both pics, even a wheat field gets blocked! :O
-
Hyperbole criticism and AI experts taking themselves too seriously
By
–
my quote tweets are filled with the fiery burning rage of people who cant take hyperbole in stride and are very proud of being “Internationally Recognized” in AI/ML as though loudly declaring this in their bio makes them any less ngmi
-
Work Culture Anxiety: How Technology Disrupts Modern Sleep
By
–
When I close my eyes, the following happens 1. I remember all my pending activities and this create guilt (tasks never ends)
2. I cannot handle the anxiety and open my phone
3. Sleep gets spoiled and I feel guilt again Why do we work like machines? #lifeisstrange -
Petri Framework Advances AI Model Safety Assessments
By
–
Petri builds on our alignment assessments in the Claude 4 and 4.5 System Cards; the @AISecurityInst also successfully built on a pre-release version of Petri for their assessments of our models.
-
Petri Open-Source AI Safety Research Tool Released
By
–
Petri is open-source and available now: http://
github.com/safety-researc
h/petri
… Read the full technical report: https://
alignment.anthropic.com/2025/petri -

Anthropic Open-Sources AI Audit Tool for Claude Sonnet 4.5
By
–
Last week we released Claude Sonnet 4.5. As part of our alignment testing, we used a new tool to run automated audits for behaviors like sycophancy and deception. Now we’re open-sourcing the tool to run those audits.