I don’t disagree. Let’s say most capable open-weight LLM when averaged over all major benchmarks (reasoning, coding, logic, tool use, math, knowledge)
Debate on most capable open-weight LLM averaged across benchmarks
By
–
By
–
I don’t disagree. Let’s say most capable open-weight LLM when averaged over all major benchmarks (reasoning, coding, logic, tool use, math, knowledge)