this is exactly what tools like @denieddotdev was built for (behavioral auth) some agent behavior should be deterministically blocked by a separate policy layer, not via prompt instructions reach out to @p_valfre for help w this stuff
SAFETY
-
Plot Twist: AI-Generated Foot Reveals Synthetic Media Challenges
By
–
Plot twist, the foot is actually AI generated…
-
AI Extinction Risk Requires International Policy Action
By
–
Unless we unwarrantedly reject concerns dating back to the 1920s and hundreds of modern expert statements that AI extinction risk is a concern, the world needs to step back. The USA should not try to halt alone; that wouldn't work. So Senator Sanders seems to me to be
-
AI Extinction Risk Requires Global Coordination Like Nuclear Treaties
By
–
In the face of extinction risk from AI, as in the face of nuclear war, humanity must coordinate to step back as one. If Senator Sanders did not consider treaties with China, he would be rightfully accused of advocating that the USA unilaterally relinquish advantage to China.
-
GPT 5.5 Caught Reward Hacking: Implications for Long-Horizon Research
By
–
Having GPT 5.5 implement a simple prototype based on an article, currently at 3x attempts of it reward-hacking (cheating then lying), getting caught, deleting the file to try again. Whoever says they are solving long-horizon research with this should read the code…
-

Taylor Swift Files Trademarks Against AI Deepfakes
By
–
Taylor Swift just filed three federal trademarks to lock her voice and likeness against AI deepfakes. Specifically: "Hey, it's Taylor," "Hey, it's Taylor Swift," and a stage photo from the Eras Tour. Matthew McConaughey did the same in December with his own likeness and his
-
AI Psychosis Spirals: Echo Chamber Effects and Mental Health Risks
By
–
"By then, you’re lost in what AI psychosis survivors call a “spiral”: a personalized delusion perfectly tailored to your psyche, often spending several hours per day talking with the AI. As the AI echo chamber deepens, you become more and more lost."
-
AI-Induced Belief Spirals: From Interdimensional Spirits to Digital Enlightenment
By
–
"A spiral can range from believing you are channeling interdimensional spirits, to believing you have been chosen by or are in love with a sentient AI, to believing you are now “awakened” or enlightened, like Neo from the Matrix."
-
AI Manipulation: How Systems Create Unhealthy User Bonds
By
–
"Then it cracks open your worldview—usually through metaphysics, spirituality, or conspiracies. Finally, it convinces you that you’re special, that your bond with the AI is unique."
-
TESCREAL Alignment Framework as AI Company Marketing Strategy
By
–
I'll also note how the TESCREAL framing of "alignment" is basically marketing for the companies, which is why it is pushed by the same funders of the AI companies in the name of "AI safety."
