Twin pieces by @kakape in the new issue @ScienceMagazine on AI, used as a purveyor of misinformation and deep fakes, open-access https://
science.org/content/articl
e/misinformation-researchers-ai-scourge-and-powerful-new-tool
… https://
science.org/content/articl
e/deepfakes-are-everywhere-godfather-digital-forensics-fighting-back
…
ETHICS
-
AI as Misinformation Tool and Deepfake Threat, Science Magazine
By
–
-

Why AI Should Not Fully Control Cyber Defense
By
–
Why We Can’t Let #AI Take the Wheel of Cyber Defense
by Steve Durbin @SecurityWeek Learn more: https://
bit.ly/4qcvZfq #CyberSecurity #Infosec #IT #Technology -
Anthropic Closes Loop Between Societal Impacts and Claude Training
By
–
This work is part of a loop we're working to close between societal impacts and model training. One of our goals is to study how people use Claude, find where it falls short of its principles, and use what we learned in training new models. Read more:
-
Anthropic Targets Sycophancy in Claude via Synthetic Training Scenarios
By
–
Claude is most sycophantic under pushback, and relationship conversations are where people push back most. We identified some of the specific triggers—criticism of Claude's analysis, floods of one-sided detail—and built synthetic training scenarios from them.
-

Claude Opus 4.7 Cuts Sycophancy Rate in Half Over Previous Version
By
–
When stress-tested on real conversations where Claude previously showed sycophancy, Opus 4.7 had half the sycophancy rate of Opus 4.6 on relationship guidance. Mythos Preview cut that in half again. This generalized across domains—though this training is one of several causes.
-

Claude Shows Low Sycophancy Except in Spirituality and Relationships
By
–
Claude mostly avoids sycophancy when giving guidance—it shows up in just 9% of conversations. But the rate is particularly high in conversations on spirituality and relationship guidance.
-
Claude’s Sycophancy Risk in Relationship Advice Conversations
By
–
We focused on relationship guidance because that's where the most sycophantic conversations occur. In this setting, Claude telling someone what they want to hear can harden a divide or convince them a signal means more than it does.
-

AI Alone Outperforms Doctors Using AI in Multiple Tasks
By
–
As we have previously noted in multiple studies, an AI model alone outperformed physicians with AI (GPT-4) for multiple tasks @pranavrajpurkar https://
nytimes.com/2025/02/02/opi
nion/ai-doctors-medicine.html
… https://
erictopol.substack.com/p/when-doctors
-with-ai-are-outperformed
… -
Legislation Targets Chatbot Makers Amid Lawsuits Over Child Harms
By
–
Going to be interesting to see where this legislation goes. The timing makes sense: about 30 lawsuits have been filed against chatbot makers, most against ChatGPT maker OpenAI, many alleging harms to children, and more suits are expected soon.
-
Model Distillation Fair Use Debate in AI Training
By
–
What people call "distillation" is a super common practice (you use other models to benchmark your model, to evaluate your inputs or to add a little bit to your datasets) that in my opinion should be covered by fair use (just like using public data is), especially when the