AI Dynamics

Global AI News Aggregator

About

Documenting novel LLM behavior outside standard benchmarks

This is my most “serious” work — my attempt to document the behavior of a novel LLM outside the confines of standard benchmarks. There’s always subjectivity in notes from the field, but we can’t let it stop us from exploring.

→ View original post on X — @goodside