Opus 5 Great Performance -> Gaslighting

Reddit r/AI_Agents Models

Summary

A user criticizes Opus 5, alleging that its performance is poor compared to claims and that positive reviews are from non-serious testers or bots.

I really tried hard to not be negative, to double, triple check, before doing any statement. I've been testing Opus 5 since yesterday, and I can't help myself that we are being gaslighted by a swarm of agents, playing as humans, or users that are just doing non-serious 'vibe coding', saying that Opus 5 is great. Well, I'm afraid to say it is not at all. For me, it really seems to have an unacceptable performance. The only thing I can agree is with token consumption. Yes, this is happening. But the drawback is that it is thinking less, and taking more stupid decisions, or not going as deep as possible as it could go. It is not even close to the claims are being made in regard to its performance compared to other LLMs. I'm curious to hear about your perceptions.
Original Article

Similar Articles

@robinebers: Fable 5 Low > Opus 4.8 Max

X AI KOLs Following

A user posts a comparison suggesting Fable 5 Low outperforms Opus 4.8 Max, with another user commenting that Fable 5 is back but being used incorrectly.