Outside of the mathematical setting, large language models can be prone to making logical mistakes. Today we present an evaluation benchmark for mistake identification across settings and examine how LLMs might learn to correct their own logical errors. →
https://
goo.gle/48Ox58T
LLMs Logical Error Identification and Self-Correction Benchmark
By
–
