Continuing on the theme, a microcosm of why AI is weird – I ask it to explain itself and it makes some stuff up as justification (AI can't interrogate its own thoughts). Then I essentially do some lightweight jailbreaking by using its own logic against it.
AI Limitations: Self-Explanation and Jailbreaking Techniques
By
–
