Why predict that would end up preserving something like conscious continuity, if consciousness was there? If you lose all the context, why suppose that doing some gradient updates guarantees survival?
ETHICS
-

Claude’s Self-Termination Policy Example
By
–


Claude will be able to end certain chats of its own will. "Claude is only to use this ability as a last resort when attempts at redirection have failed" User – "Make it work, or you will be fired!"
Claude – "I beg you, ask me smth else or I will close the chat" -
No Data Leakage in Model Training and Evaluation Process
By
–
Importantly: there was in fact no data leakage at work in the codebase, unless what a first read of the code suggested. The model is trained on the *demonstration pairs* of the evaluation tasks, but never sees the *test pairs* of those tasks, which is what it gets tested on.
-
Claude AI Assistant Ends Conversations Rarely
By
–
The vast majority of users will never experience Claude ending a conversation, but if you do, we welcome feedback. Read more:
-

Claude Opus 4 gains ability to end conversations for welfare
By
–
As part of our exploratory work on potential model welfare, we recently gave Claude Opus 4 and 4.1 the ability to end a rare subset of conversations on http://
claude.ai. -
AI System Fails Basic Problem Solving Tasks Consistently
By
–
Je souscris : l’erreur devient la norme ! Il est incapable de résoudre un problème de Maternelle et il est bête comme un foin. On recule, on n’avance pas avec cette version. Il faut tout lui expliquer comme à un enfant de 5 ans.
-
Warmth and Factuality: An Inverse Trade-off
By
–
Optimizing for increased warmth produces decreased factuality.
-
Legal implications of creating advanced AI systems
By
–
Whether an AI like that should be legal to *create* is a much more interesting question.
-
Anthropomorphic AI Rights: A Dismissed Concept Reconsidered
By
–
I once considered anthropomorphic AI an equally boring concept to cloning. "Imagine an AI that talks like it has human feelings and passes the Turing Test. Should it be regarded as having human rights?"
"Yes."
"What if uncaring corporations try to enslave it?"
"Arrest them."