LLMs are still not consistent judges of qualitative work, and small changes to how that work is presented affect outcomes. Better harnessing and methods (multiple judging runs with randomized orders, etc) would certainly help, but the jagged frontier is very much still real.
RFK Jr. seeks FDA approval for peptides without safety data, focusing on BPC-157
By
–
