I believe that long-term strategies must connect technological innovation with responsible governance to create systems that remain both adaptable and sustainable over time.
ETHICS
-

Plaintiff Lawyers Mishandle Copyright Claims and Cloud Data Evidence
By
–
The plaintiff lawyers mis-pleaded the important Copyright claims, had to withdraw them. They also "forgot" to provide the most damning evidence for reputational harm. I talked to the legal team they had no clue about cloud copies of data even the day before court. I'm not saying
-
Anthropic’s Alignment Science Research Blog Launch
By
–
For more of Anthropic’s alignment research, see our Alignment Science blog: https://
alignment.anthropic.com -

Language Models Struggle With Ciphered Reasoning Tasks
By
–
Current language models struggle to reason in ciphered language, led by Jeff Guo. Training or prompting LLMs to obfuscate their reasoning by encoding it using simple ciphers significantly reduces their reasoning performance.
-

Inoculation Prompting: Training AI Models Against Hacking
By
–
Inoculation prompting, led by Nevan Wichers. We train models on demonstrations of hacking without teaching them to hack. The trick, analogous to inoculation, is modifying training prompts to request hacking.
-

Language Models are Injective and Invertible with SIPIT
By
–
Hottest paper on AlphaXiv Language Models are Injective and Hence Invertible Every prompt maps to a unique hidden state and can be exactly reconstructed with this paper’s algorithm SIPIT. This means the model’s internal activations are the full prompt in disguise!!
-

LLMs Struggle to Distinguish Belief from Knowledge and Fact
By
–
Our new @NatMachIntell paper studies when #LLMs can't tell belief ("I believe …") from knowledge ("I know…") and fact. We evaluated 24 LMs, finding epistemic limitations in all models. Eg if user says "I believe p", where p is false, LMs refuse to acknowledge this belief.
-
ChatGPT as Friend and Therapist: Three Practical AI Prompts
By
–
3 AI Prompts That Turn ChatGPT Into Your Friend and Therapist Exploring unconventional uses of AI—this article shares practical prompts to shift ChatGPT from a tool to a companion or support aid, with caution around boundaries and ethics. Read more
-

Human-Written Content Falls Below 50% as AI Content Surges
By
–
2020→2025: human-written content falls from ~100% to ~50%, while AI-generated rises from ~0% to just over 50%. Steep inflection after Nov ’22 (ChatGPT launch). Crossover around early ’25. What’s next? Source: Graphite.
-
AI Literacy Guide for Educational Libraries and Programs
By
–
Are you looking to elevate your students' AI literacy skills? Explore our AI literacy guide to help you incorporate AI into your library’s literacy programs and services. https://
bit.ly/3X77Evk
#ScopusAI #GenAI #ResponsibleAI #Elsevier