That's so weird... This doesn't at all match my experience with Claude. I've never seen it behave this way.
Haha this is spot on how I've been feeling lately. I find it unbearable to work with this model for this reason... any trick out there you can do to steer it not to overcomplicate things? I guess Codex here I come
This is actually what keeps people using AI: variable reward schedule. It's basically gambling.
I was expecting it to spend 30 minutes running headless chrome instances, taking screenshots and analyzing them in python to verify the blueness of the result.
I never really understood what being "triggered" was like until now.
I lost it when it finally did the right thing, but then it added a never-requested gradient to the button.
Very good!
> 23 agents total.<p>This hit a bit too close to home. Sol has the same issue, spawns a lot of agents for no good reasons (besides burning tokens).
Brilliant. Precisely the reason I stopped using Anthropic's products.
This is so good at replicating the experience of frustration, then relief when it finally does what you asked it to do in the first place!
I find it funny how it went off with subagents and adversarial review when a simple grep or diff is sufficient.
Congratulations!!! You win what’s left of the internet - just ask Claude for your prize! Motrin I’ve had this week.<p>I use codex now.
I don't get the joke... maybe because I'm using Codex?
It's funny but unrealistic as Claude does a pretty good job at only changing what is required these days with the 5 tier models like Opus 5 or Fable.
The site is opusfived.com, and Opus 5 is probably the worst so far at doing this.
I dunno, in my experience it's often overly-narrow, sometimes jumping through all kinds of hoops to preserve some edge-case behaviour that doesn't matter because I didn't mention it could be changed.
This is exactly the experience I have with Opus 5. Opus 4.6 is better, Flable 5.1 much better. But Opus 5 is infuriating.
I've experienced this so many times over.<p>"I was wrong" and "the honest truth" are just forever phrases that are now dead to me.
How do you manage your frustration in these interactions? I often find myself getting pissed off
I stop using AI and do the job manually. I normally give AI one shot at the task. If it fails then it's not saving me any time
The goal of my personal harness is to get to the point where I never actually talk to Claude directly for that very reason.
I’m laughing and crying at the same time. This is what work feels like now. Thank you, well done!
This is so perfect and depressing that I might cry. It’s like a Kafka novel about programming.
This is pure genius. No notes
So this site is just a fan-fiction that thinks it's somehow dunking on Claude? I've never had a session that remotely resembles any of this. I honestly can't tell what point this site thinks it's making.
I wrote my own harness to stop shit like this from getting to my attention out of frustration.<p>I'm sure there's quite a bit of variation from person to person in these sorts of experiences, based on your harness, the way you talk, the stored memory, your CLAUDE.md, etc. But people <i>absolutely</i> have had this Opus 5 style experience the app simulates.
So you are lucky, congratulations.
[flagged]
<a href="https://www.scribbr.com/fallacies/either-or-fallacy/" rel="nofollow">https://www.scribbr.com/fallacies/either-or-fallacy/</a>
The website is literally built with Lovable.dev what are you talking about
This is scary close to my interaction with Claude this week.
This spiked my blood pressure. Well done
Am I the only one whose experience doesn't match this?<p>My gripe with Claude is that while investigating how to do this it will report 200 other incidental findings which I overlooked and I realize those are broken too and need urgent fixing, derailing <i>me</i>, not <i>it</i>.
Oh man, exactly. I'm very prone to scope creep as I work on tasks. I already would notice some things that could be fixed or refactored and have a hard time not touching them before I used agents. But now I have to be very intentional about not letting it manipulate me into fixing EVERYTHING RIGHT NOW. Half the time the "one more thing worth noting, unrelated..." isn't even an actual issue, it just brought it up to fish more usage out of me.<p>Also, while this little demo is certainly exaggerating the issue, I do find working with Claude to sometimes get quite verbose and tiresome. I doubt I would struggle this much to get it to change a button color, but the patterns of speech, the endless lists, the over-explanations, and the whole song and dance of trying to get it to make the change you want without side-effects is frustratingly familiar to me.
Claude is an unbelievable yak shaver if you let it be.
You're not alone. I've been sat wondering what kind of codebase someone has if they have this problem, I've <i>never</i> seen this behaviour.
Just when I came back to my pc and was thinking "I hate this world were everyone talks about AI like fanatics" this made me a little bit happy, especially the unhingend all caps options towards the end
Here come all the totally organic "wow, I guess I better switch to OpenAI" comments.
now THAT is a load-bearing simulation
PTSD 9000.. I miss the old days, less load bearing BS and more in the zone coding..
Thanks, I hate it.
rofl, brilliant
Lol this is great
[dead]