AI Dynamics

Global AI News Aggregator

About

RLHF Fine-tuning Effects on Model Behavior Analysis

That example looks to me like it's more caused by RLHF fine-tuning than anything that was baked into the model in the pre-training phase

→ View original post on X — @simonw