Working on self-contained ones as separate posts but a general one: I work with content moderation datasets. I often send big CSV of offensive filth to an LLM to annotate and get back refusals or quasi-refusals. Any competent "unhinged" fallback LLM saves me work.
Unhinged LLM fallback for content moderation annotation
By
–