AI Dynamics

Global AI News Aggregator

About

Evidence Levels and n=1 Experimentation in AI Evaluation

You seem over-indexing on "Everything here could have been done without ChatGPT". But that's the least interesting bit. We don't "need to prove this before anyone can believe" – there are levels of evidence short of proof. We don't need "clear, resounding successes" to try n=1

→ View original post on X — @jeremyphoward