I have little to say about GPT-5.x for the same reason I’ve long had little to say about Claude 4.x—which is that every txt2txt LLM task is at most two: 1) A task with real economic value
2) A task that differentiates frontier models
3) A task you would gladly read in a tweet
Three categories of LLM tasks: economic, differentiating, and tweetable
By
–