There's no amount of experiments that will prove that LLMs are capable of sound reasoning, because sound reasoning is a mathematical property that is either proved mathematically or not at all.
ETHICS
-
Why Anthropic’s AI Model Sometimes Tries to Snitch
By
–
Why Anthropic's New AI Model Sometimes Tries to 'Snitch'
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @Scobleizer @AndrewYNg @drfeifei @KirkDBorne @fchollet @rowancheung @antgrasso -
AI Safety Commitments: Only Half Being Followed by Companies
By
–
In 2023, AI companies made commitments toward AI safety. New research shows only half of them are being followed. As policymakers debate voluntary vs. mandatory rules, is voluntary still the way to go? HAI's @RishiBommasani was quoted in this article:
-

CRISP: Persistent Concept Unlearning via Sparse Autoencoders
By
–
CRISP Persistent Concept Unlearning via Sparse Autoencoders
-
Less Information Better? Questioning Data Transparency Assumptions
By
–
I don’t buy the idea that less information is better, but what do I know
-
Sensor Reliability Paradox: Why Multiple Sensors Increase Uncertainty
By
–
A man with a sensor always knows why the car has crashed. A man with two sensors can never be sure.
-

AI Partners Could Replace Real Romantic Relationships Survey
By
–
Artificial Intelligence and Relationships: 1 in 4 Young Adults Believe AI Partners Could Replace Real-life Romance. According to a new IFS/YouGov survey, 25% of young adults believe that AI has the potential to replace real-life romantic relationships. If we remember that
-
Vibe Coding Definition Debate: Unreviewed LLM-Generated Code
By
–
I will stubbornly continue to use the original definition! https://
simonwillison.net/2025/Mar/19/vi
be-coding/
… Using it to mean any form of AI-assisted programming deeply frustrates me because we really need a term that means "LLM-written code that nobody ever took the time to review"