

It's crazy that a 4B model can match GPT-OSS-120B **even on knowledge benchs** like MMLU and GPQA! Knowledge was long a limitation of smol models ; not anymore. This 4B could be a good candidate for the "agentic core" model described by @karpathy
By
–



It's crazy that a 4B model can match GPT-OSS-120B **even on knowledge benchs** like MMLU and GPQA! Knowledge was long a limitation of smol models ; not anymore. This 4B could be a good candidate for the "agentic core" model described by @karpathy