AI Dynamics

Global AI News Aggregator

About

The Ambiguity Problem in LLM Labeling and Training Data

Consider being a labeler for an LLM. The prompt is “give me a random number between 1 and 10”. What SFT & RM labels do you contribute? What does this do the network when trained on? In subtle way this problem is present in every prompt that does not have a single unique answer.

→ View original post on X — @karpathy