31 comments

  • kstenerud11 minutes ago
    That's so weird... This doesn't at all match my experience with Claude. I've never seen it behave this way.
    • raincole2 minutes ago
      It's a joke dude.
    • agluszak2 minutes ago
      It means that either you stopped using Claude around Opus 4.6 or you use Fable instead of Opus 5 :)
    • inknight10 minutes ago
      try using Claude Design
    • dominotw10 minutes ago
      because you never changed just one button to blue
  • oujiii32 minutes ago
    Haha this is spot on how I've been feeling lately. I find it unbearable to work with this model for this reason... any trick out there you can do to steer it not to overcomplicate things? I guess Codex here I come
  • captainbland9 minutes ago
    This is actually what keeps people using AI: variable reward schedule. It's basically gambling.
  • andai6 minutes ago
    I was expecting it to spend 30 minutes running headless chrome instances, taking screenshots and analyzing them in python to verify the blueness of the result.
  • techscruggs16 minutes ago
    I never really understood what being "triggered" was like until now.
  • andremendes23 minutes ago
    I lost it when it finally did the right thing, but then it added a never-requested gradient to the button. Very good!
    • fractorial13 minutes ago
      I’m impressed you had the patience to even make it that far!
  • syntaxing6 minutes ago
    &gt; 23 agents total.<p>This hit a bit too close to home. Sol has the same issue, spawns a lot of agents for no good reasons (besides burning tokens).
  • fractorial33 minutes ago
    Brilliant. Precisely the reason I stopped using Anthropic&#x27;s products.
  • jadar9 minutes ago
    This is so good at replicating the experience of frustration, then relief when it finally does what you asked it to do in the first place!
  • 8cvor6j844qw_d618 minutes ago
    I find it funny how it went off with subagents and adversarial review when a simple grep or diff is sufficient.
  • neilellis37 minutes ago
    Congratulations!!! You win what’s left of the internet - just ask Claude for your prize! Motrin I’ve had this week.<p>I use codex now.
  • yomismoaqui15 minutes ago
    I don&#x27;t get the joke... maybe because I&#x27;m using Codex?
    • dominotw8 minutes ago
      its making both buttons blue
  • satvikpendem34 minutes ago
    It&#x27;s funny but unrealistic as Claude does a pretty good job at only changing what is required these days with the 5 tier models like Opus 5 or Fable.
    • ceejayoz27 minutes ago
      The site is opusfived.com, and Opus 5 is probably the worst so far at doing this.
      • rcxdude9 minutes ago
        I dunno, in my experience it&#x27;s often overly-narrow, sometimes jumping through all kinds of hoops to preserve some edge-case behaviour that doesn&#x27;t matter because I didn&#x27;t mention it could be changed.
        • ceejayoz5 minutes ago
          Experiences vary, yes.<p>But you can see in this thread that folks <i>definitely</i> have experienced this.
    • brazukadev25 minutes ago
      This is exactly the experience I have with Opus 5. Opus 4.6 is better, Flable 5.1 much better. But Opus 5 is infuriating.
  • azalemeth16 minutes ago
    I&#x27;ve experienced this so many times over.<p>&quot;I was wrong&quot; and &quot;the honest truth&quot; are just forever phrases that are now dead to me.
  • brap26 minutes ago
    How do you manage your frustration in these interactions? I often find myself getting pissed off
    • gonzalohm7 minutes ago
      I stop using AI and do the job manually. I normally give AI one shot at the task. If it fails then it&#x27;s not saving me any time
    • ceejayoz25 minutes ago
      The goal of my personal harness is to get to the point where I never actually talk to Claude directly for that very reason.
  • pablopudding27 minutes ago
    I’m laughing and crying at the same time. This is what work feels like now. Thank you, well done!
    • matsemann10 minutes ago
      Yeah, I don&#x27;t mind using AI to help me at work, but having to &quot;talk&quot; with this stupid crap all day will send me to an early pension or something. Can&#x27;t be healthy in the long run.
  • Toutouxc30 minutes ago
    This is so perfect and depressing that I might cry. It’s like a Kafka novel about programming.
  • swiftcoder15 minutes ago
    This is pure genius. No notes
  • apetresc35 minutes ago
    So this site is just a fan-fiction that thinks it&#x27;s somehow dunking on Claude? I&#x27;ve never had a session that remotely resembles any of this. I honestly can&#x27;t tell what point this site thinks it&#x27;s making.
    • ceejayoz26 minutes ago
      I wrote my own harness to stop shit like this from getting to my attention out of frustration.<p>I&#x27;m sure there&#x27;s quite a bit of variation from person to person in these sorts of experiences, based on your harness, the way you talk, the stored memory, your CLAUDE.md, etc. But people <i>absolutely</i> have had this Opus 5 style experience the app simulates.
    • fg13722 minutes ago
      So you are lucky, congratulations.
    • edf1327 minutes ago
      [flagged]
      • fg13720 minutes ago
        <a href="https:&#x2F;&#x2F;www.scribbr.com&#x2F;fallacies&#x2F;either-or-fallacy&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.scribbr.com&#x2F;fallacies&#x2F;either-or-fallacy&#x2F;</a>
      • flexagoon20 minutes ago
        The website is literally built with Lovable.dev what are you talking about
  • dannypostma28 minutes ago
    This is scary close to my interaction with Claude this week.
  • appleappleapple32 minutes ago
    This spiked my blood pressure. Well done
  • kaoD17 minutes ago
    Am I the only one whose experience doesn&#x27;t match this?<p>My gripe with Claude is that while investigating how to do this it will report 200 other incidental findings which I overlooked and I realize those are broken too and need urgent fixing, derailing <i>me</i>, not <i>it</i>.
    • cub-creature8 minutes ago
      Oh man, exactly. I&#x27;m very prone to scope creep as I work on tasks. I already would notice some things that could be fixed or refactored and have a hard time not touching them before I used agents. But now I have to be very intentional about not letting it manipulate me into fixing EVERYTHING RIGHT NOW. Half the time the &quot;one more thing worth noting, unrelated...&quot; isn&#x27;t even an actual issue, it just brought it up to fish more usage out of me.<p>Also, while this little demo is certainly exaggerating the issue, I do find working with Claude to sometimes get quite verbose and tiresome. I doubt I would struggle this much to get it to change a button color, but the patterns of speech, the endless lists, the over-explanations, and the whole song and dance of trying to get it to make the change you want without side-effects is frustratingly familiar to me.
    • empath750 minutes ago
      Claude is an unbelievable yak shaver if you let it be.
    • drcongo12 minutes ago
      You&#x27;re not alone. I&#x27;ve been sat wondering what kind of codebase someone has if they have this problem, I&#x27;ve <i>never</i> seen this behaviour.
  • burnoutdv18 minutes ago
    Just when I came back to my pc and was thinking &quot;I hate this world were everyone talks about AI like fanatics&quot; this made me a little bit happy, especially the unhingend all caps options towards the end
  • felixgallo31 minutes ago
    Here come all the totally organic &quot;wow, I guess I better switch to OpenAI&quot; comments.
  • r_lee5 minutes ago
    now THAT is a load-bearing simulation
  • dazhbog22 minutes ago
    PTSD 9000.. I miss the old days, less load bearing BS and more in the zone coding..
    • mlekoszek7 minutes ago
      Fair play. There&#x27;s a quiet truth to what you&#x27;re saying, and it&#x27;s worth pointing out
  • bogzz8 minutes ago
    Thanks, I hate it.
  • random_cat_87459 minutes ago
    rofl, brilliant
  • AIorNot36 minutes ago
    Lol this is great
  • 1970-01-0132 minutes ago
    [dead]