AI Dynamics

Global AI News Aggregator

About

AI Model Behavior and RLHF Tuning in ChatGPT

Yeah this got fixed fast in ChatGPT. It was never that susceptible to it anyway — the RLHF-tuned models don't go wild like the original reports show on davinci-instruct-beta

→ View original post on X — @goodside