cyberpsychosis is real now. rogue ai's are real now. who's building arasaka?
SAFETY
-
Anthropic’s Unique Position Among AI Frontier Labs on Alignment
By
–
i was just thinking that anthro is the only frontier lab that hasnt had a major alignment embarrassment – v fitting
-

Software Engineers Unafraid of AI Tool Obsolescence
By
–
I continue to be entirely unafraid that these tools are going to obsolete my skills as a software engineer
-
EU Commission Launches Age Verification App Prototype for Child Safety
By
–
Joint press release: Commission presents guidelines and age verification app prototype for a safer online space for children https://
ec.europa.eu/commission/pre
sscorner/detail/en/ip_25_1820
…
#digitaleu #online #SAFE #Innovation #technology @DigitalEU @elaniazito @ArturHabant @BetaMoroney @AkwyZ @AnthonyRochand -
LLMs Tendency to Generate Excessive SCP Style Text
By
–
Here's an example from over a year ago showing how LLMs just *love* to produce token after token of text in the SCP style!:https://t.co/EUocjpdECI
— Jeremy Howard (@jeremyphoward) 17 juillet 2025Here's an example from over a year ago showing how LLMs just *love* to produce token after token of text in the SCP style!:
-
ChatGPT-Induced Psychosis Emerges as Active Research Field
By
–
There are, sadly, now enough examples of chatgpt-induced psychosis that it's an active field of study.
-
Monitoring AI Services for User Safety and Abuse Detection
By
–
I'm not sure the best way to counter this. Perhaps services can use the monitoring layer then nearly all use to look for copyright violations, system prompt hacks, etc, to also look for signs a user may be taking a role play too seriously, and let them know they're just playing?
-

Psychiatrists warn chatbots may trigger psychosis risk
By
–
Psychiatrists have been warning about the potential for chatbots to trigger psychosis for some years. https://
pmc.ncbi.nlm.nih.gov/articles/PMC10
686326/
… -
ChatGPT’s Self-Reinforcing Distribution Loop and Memory Issues
By
–
This created a self-reinforcing feedback loop. The more in-distribution tokens ChatGPT was getting in its chat history, the more strongly the auto-regressive model was pushed to stay in that distribution. ChatGPT memory made this even worse, letting it happen across chats.
-
AI Agents Financial Security Risks and Transaction Limitations
By
–
You can have it enter CC information on your behalf if you want. So it could potentially send money to scams, etc. That was the example they gave. Probably best to have it hand back off to you for any transactions in the short-term.