i'm a few days late to realizing this but: wow, opus 4.7 is god awful like so, so bad it's making mistakes on things i'd expect gpt-4o to handle cleanly there's got to be some explanation, right?
Opus 4.7 Performance Issues: Major Regression Compared to GPT-4o
By
–