
I re-ran an experiment @colin_fraser ran against GPT-4o a while back to see how good it was at adding long numbers, only this time I tried Qwen 3.8 27B running locally in both reasoning and non-reasoning modes https://
simonwillison.net/2026/Oct/4/qwe
n38-addition-in-words/
…
