AI Dynamics

Global AI News Aggregator

About

How RLHF Shapes ChatGPT’s Helpful and Friendly Tone

ChatGPT doesn't sound helpful by accident. Someone taught it to sound that way. After the model learns to follow instructions, labs show it thousands of answer pairs.
Humans pick which response feels clearer, friendlier, or more useful. The model learns the preferred style.

→ View original post on X — @whats_ai