The base LLM tried to solve math problems like “a^3 · a^2 if a=5,” where it incorrectly calculated the answer to be 125. Co-LLM trained the model to collaborate w/the large math LLM Llemma, and together they determined that the correct solution was 3,125.
Co-LLM Collaboration Improves Mathematical Problem Solving Accuracy
By
–