AI Dynamics

Global AI News Aggregator

About

Custom Prompts and Benchmark Evaluation Standards for LLMs

You mention our "custom prompt" like if there was an official way of prompting (yours?). Most benchmarks were created before this concept of LLM eval with prompting even exists, and for many of them there is no official prompt or way to evaluate them with LLMs.

→ View original post on X — @guillaumelample