Why Anthropic's New AI Model Sometimes Tries to 'Snitch'
#AI #AIio #AIInnovation #ML #DataScience #Futureofwork @Scobleizer @AndrewYNg @drfeifei @KirkDBorne @fchollet @rowancheung @antgrasso
SAFETY
-
Why Anthropic’s AI Model Sometimes Tries to Snitch
By
–
-
Pedagogy in the Face of AI: The Black Box
By
–
Study link: https://researchgate.net/publication/385922814_Learning_to_work_with_the_black_box_Pedagogy_for_a_world_with_artificial_intelligence
… -
Trusting AI: When and How to Judge It
By
–
So the challenge isn’t “make AI explainable.” The challenge is: → When should you trust AI?
→ When should you doubt it?
→ How do you judge quality in a system you can’t dissect? -

The Truth About AI’s Black Box
By
–
Everyone talks about "understanding AI." But here’s the truth: You can’t open the black box. You can only work around it. Even the engineers can’t explain why a model gave you that output.
-
Darwin Gödel Machine: Novelty and Implications Commentary
By
–
Darwin Gödel Machine: A Commentary on Novelty and Implications, from @AntoMon https://
antomon.github.io/posts/darwin-g
odel-machine/
… -
Waymo Driverless Taxis Targeted by Protesters in Los Angeles
By
–
Waymo Driverless Taxis Become Protesters’ New Favorite Target Robot-operated electric vehicles were summoned to downtown Los Angeles and set alight
#RiseoftheRobots https://
wsj.com/us-news/waymo-
driverless-taxis-become-protesters-new-favorite-target-23405f0a?st=Qbw4D6
… via @WSJ -

Pseudo-Simulation: New Autonomous Driving Evaluation Paradigm
By
–
Pseudo-Simulation for Autonomous Driving Pseudo-simulation is a new evaluation paradigm for autonomous vehicles that blends the realism of real-world data with the generalization power of simulation, enabling robust, scalable testing without the need for interactive
-

Paper IA: Demand for Reasoning Traces Analysis and Transparency
By
–
Y estos tweets de @scaling01 y todo el análisis que ha compartido es también demoledor. Estaría genial que se publicaran las trazas de razonamiento sobre las que se sustenta el paper para analizarlas bien.
-

Reasoning Models Limitations: Understanding Problem Complexity Trade-offs
By
–
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
Paper: https://
ml-site.cdn-apple.com/papers/the-ill
usion-of-thinking.pdf
… -

Can AI truly understand humor and jokes?
By
–
Has AI ever understood a joke? PS: These twitter slop generators really piss me off. Wtf is the point of them?