
Fantastic work — imagine combining this with “LLMs Can Self-Improve”! Btw, I suggested LLM-inferred consensus in a tweet last year — AFAIK I thought of it from reading Xuezhi Wang et al. but there may be earlier examples. Case study in “Ideas are cheap; verification is hard.”
