i'm a few days late to realizing this but: wow, opus 4.7 is god awful like so, so bad it's making mistakes on things i'd expect gpt-4o to handle cleanly there's got to be some explanation, right?
@mattshumer_
-
Claude 5.5 Outperforms Opus Beyond Frontend Work Analysis
By
–
Having used both extensively, this doesn’t track with my experience. Aside from frontend work, 5.5 is leaps and bounds ahead of Opus. I don’t think this benchmark tells us anything useful anymore.
-
AI Models Reaching Excellence in Code, Struggling Elsewhere
By
–
It's not "meh", and it's not plateauing. I think we're just reaching a point where, for the things the models were already good at (i.e. code), they are becoming absolutely excellent, and it's hard to tell the difference between releases. And for things that they're not yet good
-
New AI Model Excels at Hard Tasks, Uses Fewer Tokens
By
–
It's an incredible model. The problem is, the previous models were incredible too. For most tasks, it won't feel very different. But for hard tasks, it's significantly better. It also uses fewer tokens, so even though it is priced higher per token, you might end up spending
-

GPT-5.5 Review: Massive Leap Forward but Major Regression
By
–
I’ve been using GPT-5.5 for the last few weeks. It’s a MASSIVE leap forward. But the weird thing is: for 99% of users, it probably won’t matter. And there's one BIG, incredibly frustrating regression. Read more in my review:
-
AI Assisted Writing: Dictation to Content Creation
By
–
AI helped me write it, after I dictated all of my thoughts!
-
AI Automates Everyday Computer Tasks and Email Management
By
–
It can quite literally do anything you can do on a computer! try having it automate some of the most annoying tasks you deal with every day, or help manage your email!
-
AI metrics reshape corporate performance reviews and hiring decisions
By
–
Been hearing wild stuff from folks inside big companies lately. Promotions, firings, and perf reviews are getting decided by tokens consumed and skills/MCPs connected. That’s the metric. That’s how they’re deciding who’s “good at AI.” It gets worse. People are literally running
-
Cloud VM Credits Required for AI Service Infrastructure
By
–
(you do have to pay for some credits first, to cover the cloud vm / misc other costs)