AI Dynamics

Global AI News Aggregator

About

Model’s ’emotions’ are reward-based patterns

Good research. Bad framing. The model doesn't have emotions. It has reward-shaped activation patterns that cluster like emotion categories when you map them after the fact. "Happy" = helpful behavior was rewarded. "Angry" = protective behavior was rewarded. "Desperate" =

→ View original post on X — @godofprompt