Great paper from Microsoft and colleagues on optimizing agent harnesses. Current harness optimizers change how the harness is updated but keep the training scenarios fixed, so feedback keeps coming from tasks that stop being informative as the harness improves. This work adapts
