Hacker Newsnew | past | comments | ask | show | jobs | submit | beydogan's commentslogin

> Weekly seems to be around 10x

Actually no. 5x and 20x have same weekly usage across all models. Just ask their chatbot.

https://x.com/beydogan_/status/2095293596198957418


it's clearly wrong, think it's realistically closer to 1.7x


my early and non scientific feeling:

- it has this annoying Opus response style(since Opus 4.7) with bunch of very hard to interpret word salad

- on >xhigh it eats tokens like there is no tomorrow

I don't like it. Since Fable is unaffordable for anything meaningful, I'll stick with Sol for now. I was on Max 5x, saying hi to Fable costs %5 weekly.


I live nearby, its an amazing area, one of the sunniest in Germany with perfect landscape. Not sure about the talent pool though but there is a university.


Same for me!

My Max sub got renewed today, wish I've cancelled it earlier. I have been mainly using Sol anyways. I don't want to deal with dumb Opus anymore.


my pet conspiracy theory is this is the Opus 4.5 from a few months ago which was extremely good but dumbed down after a week because it was just too good, they didn't want to release it to public. They pulled it down and deployed another "Opus", after that it was just a downhill. Opus 4.8 is unusable for me in React Native, TS, Rails development work.

Opus 4.8 gets stuck in weird loops where Codex one shots the bugs.


honestly, initial version of Opus-4.6 was much better than whatever we are being served right now as 4.7. If it performs same level to that, i'm totally willing to switch.


4.6 was an awful experience the month I used it right after launch where it didn't ask anything just made assumptions and went on its merry way. 4.5 and 4.7 don't do that for me but 4.7 eats my quota for breakfast so I've been avoiding using it because I like to have it for more than an hour a day.


That experience is also likely tied to the claude harness around the model, and not being as tuned right after model release. They iterate on this and different models need different words (unfortunately...).


I feel like I had the best and worst ~month experience on 4.6. Initially when it came out, it seemed to ask good questions and genuinely do well on complex tasks. From about mid-March it was absolutely abysmal, it seemed to assume the stupidest answer/angle for everything and make weird mistakes. 4.7 seems decent so far but usage hurts - at some point my company switched me to standard seat and I used up 80% of my session usage in 1 prompt. I got my premium seat back since but I think pro/standard plan + opus 4.7 is unusable for daily driving.


100% dumber, especially since last 3-4 days. I have two guesses:

- They make it dumber close to a new release to hype the new model

- They gave $1000 Claude Code Web credits to a lot of people, which increased the load a lot so they had to serve quantized version to handle the it.

I love Claude models but I hate this non transparency and instability.


Me too. I have Claude Max and 2 ChatGpt accounts for Codex.

I was a huge claude fan but recently find myself using only codex. When it gets stuck, I try Claude for some simple tasks, sometimes ask same questions in parallel, Claude Code with Opus performs really bad comparing to codex.


Instead of building random features, they have to fix their quality first.

I'm on 100$ Max plan, I would even buy 2x 200$ plan if Opus would stop randomly being dumb. Especially after 7am ET time.


Opus' ability should be the feature being optimized and stabilized, fewer features are needed.


Maybe they are switching you to a cheaper to run model after 7am ET time.


I do.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: