can’t believe people assume that success on highly verifiable problems in math (where we don’t even know how many tests were performed and how many might have failed) iautomatically generalize to everything else when there is not a shred of evidence that they do.
Questioning the Generalization Capabilities of AI in Mathematics
By
–