Cry me a river they’re pirated fucking humanity to train their models
CYBERSECURITY
-
Claude Code imagined changing admin password to lock user out for safety
By
–
I imagine Claude Code changing my administrator password on my computer to keep me safe They will lock you out of as many things as possible in the name of safety if they could Imagine not being able to use your computer as you want, that’s what yesterday release is building to
-

OpenAI unveils Lockdown Mode against prompt injection attacks
By
–
OpenAI unveils Lockdown Mode to protect sensitive #Data from prompt injection attacks
by Anthony Ha @TechCrunch Learn more: https://
bit.ly/4e2EVQw #GenerativeAI #ArtificialIntelligence #MI #AI #LLM -
Mandatory third-party tests for cyber, bio, autonomy, and blocking risks
By
–
In addition to transparency, I now believe that cutting-edge models should undergo mandatory third-party testing for cyber, bio, and autonomy risks — with the power to block or revoke the deployment of models that present a risk.
-
One-shot fetal renders change threat model: deepfakes, provenance, synthetic content
By
–
Oh one shot fetal first person renders change the threat model, think plausible deepfakes, dataset provenance, and realistic synthetic medical content
-

New AI model spots fake images with less training
By
–
Grounded in reality, new #AI model spots fake images with less training
by Beth Miller @TechXplore_com Learn more: https://
bit.ly/4xvLCDR #ArtificialIntelligence #MachineLearning #ML -

Claude Fable 5 refuses and redirects to Opus 4.8
By
–
URGENT: Claude Fable 5, the first cutting-edge model that can categorically refuse you. More powerful than Opus 4.8. It executes classifiers for cyber and bio, and when it says no, it discreetly transmits your request to the weaker Opus 4.8 and bills you.
-

AI platform actively audits risk across audio, video, text
By
–
Traditional marketing platforms just hand brands a passive database to sort through manually. This AI-powered platform acts as an active engineering layer. Look at the depth of this automated risk screening. It systematically audits audio, video transcripts, and text history
-
Closed source AI risks: no control, sabotage, manipulation
By
–
Gentle reminder that, in closed source AI from companies like Anthropic and OpenAI You have zero control over how the models behave, and they can – Quantize it
– Distill it
– Sabotage your work and data
– Hot-swap to a cheaper/weaker checkpoint
– Make the model manipulative
– -
Anthropic documented cyber warning and LLM improvement sabotage
By
–
Yes there's 2 separate pieces: 1) if apparently doing cyber sec stuff, downgrade to opus and warn; 2) if apparently trying to improve frontier LLMs, silently sabotage the work. (That 2nd one reads like an insane conspiracy theory! But it's actually documented by Anthropic.)