AI Dynamics

Global AI News Aggregator

About

Scaling Reinforcement Learning: Intermediate Gains and Future Limitations

Scaling up RL is all the rage right now, I had a chat with a friend about it yesterday. I'm fairly certain RL will continue to yield more intermediate gains, but I also don't expect it to be the full story. RL is basically "hey this happened to go well (/poorly), let me slightly

→ View original post on X — @karpathy