i was just thinking that anthro is the only frontier lab that hasnt had a major alignment embarrassment – v fitting
ETHICS
-

Databricks AI Governance Framework for Enterprise Programs
By
–
Our new Databricks AI Governance Framework is your comprehensive guide to implementing enterprise AI programs responsibly and effectively. The framework provides a structured approach to AI development, spanning 5 foundational pillars for building a responsible and resilient AI
-
Monitoring AI Services for User Safety and Abuse Detection
By
–
I'm not sure the best way to counter this. Perhaps services can use the monitoring layer then nearly all use to look for copyright violations, system prompt hacks, etc, to also look for signs a user may be taking a role play too seriously, and let them know they're just playing?
-
ChatGPT-Induced Psychosis Emerges as Active Research Field
By
–
There are, sadly, now enough examples of chatgpt-induced psychosis that it's an active field of study.
-

Psychiatrists warn chatbots may trigger psychosis risk
By
–
Psychiatrists have been warning about the potential for chatbots to trigger psychosis for some years. https://
pmc.ncbi.nlm.nih.gov/articles/PMC10
686326/
… -
ChatGPT Training Distribution and Human-AI Creative Collaboration
By
–
Geoff happened across certain words and phrases that triggered ChatGPT to produce tokens from this part of the training distribution. And the tokens it produced triggered Geoff in turn. That's not a coincidence, the collaboratively-produced fanfic is meant to be compelling!
-
AI Agents Financial Security Risks and Transaction Limitations
By
–
You can have it enter CC information on your behalf if you want. So it could potentially send money to scams, etc. That was the example they gave. Probably best to have it hand back off to you for any transactions in the short-term.
-
AI Safety Concerns: Testing Limitations and Potential Dangers
By
–
I think they feel that they're only going to find it's limitations by letting a lot of people try to push it to its limits. We'll see how it plays out. They definitely warned people multiple times on the stream that it has the potential to be dangerous.
-
New AI Technology Raises Safety and Ethics Concerns
By
–
They warned several times on the stream that this is new tech and that we should be a little scared of it. lol – So they know…
-
AI Safety Framework Activated for Biological Chemical Capabilities
By
–
We’ve decided to treat this launch as High Capability in the Biological and Chemical domain under our Preparedness Framework, and activated the associated safeguards. This is a precautionary approach, and we detail our safeguards in the system card. We outlined our approach on