Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
Paper: https://
arxiv.org/pdf/2506.01939
.pdf
…
Project: https://
shenzhi-wang.github.io/high-entropy-m
inority-tokens-rlvr/
…
High-Entropy Minority Tokens Improve LLM Reasoning via Reinforcement Learning
By
–
