for the open reasoning models, response quality stayed the same not "slightly degraded." same. and here's the number that stops you: omitting assistant-side history reduced cumulative context lengths by up to 10x 10x less context. same quality. on real conversations.
10x Less Context, Same Quality in Open Reasoning Models
By
–