AI Dynamics

Global AI News Aggregator

About

AI Limitations: Self-Explanation and Jailbreaking Techniques

Continuing on the theme, a microcosm of why AI is weird – I ask it to explain itself and it makes some stuff up as justification (AI can't interrogate its own thoughts). Then I essentially do some lightweight jailbreaking by using its own logic against it.

→ View original post on X — @emollick