5 comments

  • rich_sasha1 hour ago
    I&#x27;ve noticed something else - as Anthropic models get even more and more superhuman, they seem to serve me more and more casual nonsense.<p>Not like adding glue to pizza. Here&#x27;s an example from today (paraphrasing): &quot;you need to run `git merge-base branch1 branch2`. Pay attention to the order of arguments, it is important: `git merge-base` is symmetric and returns the same value regardless of the order of inputs&quot;.<p>So which one is it? Symmetric or not? It&#x27;s not even one of those where it self-corrects, it just happily contradicts itself halfway through the sentence.<p>My pet theory is that these frontier models do quite a bit of brute force at the end - some kind of beam search - and silently downgrade you depending on demand or compute availability.<p>Still not great. This particular nugget is from Sonnet 5, default settings.
    • wongarsu20 minutes ago
      I&#x27;ve notice the same pattern even in GLM5.2. It has always been a thing, but it seems to be getting worse<p>It does feel like the kind of thing beam search would fix. It starts the sentence with a claim like &quot;Pay attention to the order of arguments&quot;. Around that time it &quot;notices&quot; that the order doesn&#x27;t matter, but it&#x27;s already committed to the sentence and has to complete it in the best way still possible<p>Maybe at some point someone figures out how to train models with a backspace token
    • sunaookami1 hour ago
      I also see stuff like this with Opus (4.8 and 5). It does the work correctly but then the explanation or &quot;aftermath&quot; description is kinda nonsense. It feels like they are tuned hard for agentic coding and everything else is getting worse. There are people that say older versions are also better for creative writing which would confirm this theory.
      • Culonavirus29 minutes ago
        &gt; they are tuned hard for agentic coding<p>Well you can&#x27;t blame them, it&#x27;s what brings in the cash. Other uses will suffer proportionally.
    • brookst20 minutes ago
      I don’t find Sonnet usable at all. Especially for the price. I’d much rather get half the usage of Opus.
    • o1044936622 minutes ago
      Claude has always noticeably degraded under heavy load. Opus goes from &quot;decent to work with&quot; to &quot;dumb intern&quot; depending on whether you&#x27;re working at 3 AM west coast or 10 AM - 5 PM. It&#x27;s part of why I cancelled my subscription - &quot;max&quot; plans and &quot;extra high&quot; effort are meaningless when there&#x27;s so much variability between model availability, model performance, and harness bugs every day and every week.
  • croemer1 hour ago
    Incident is claimed to be resolved but I&#x27;m now (11:22 UTC) getting:<p><pre><code> API Error: 529 Overloaded. This is a server-side issue, usually temporary — try again in a moment. If it persists, check https:&#x2F;&#x2F;status.claude.com. </code></pre> And the status page says all green.<p>Update 11:27 UTC: I saw the error first at 11:22 UTC. Retry at 11:27 UTC still failing. Status page is still green.<p>Update 11:28 UTC: Incident has been declared dated 11:27 UTC <a href="https:&#x2F;&#x2F;status.claude.com&#x2F;incidents&#x2F;mfdtrknpxghq" rel="nofollow">https:&#x2F;&#x2F;status.claude.com&#x2F;incidents&#x2F;mfdtrknpxghq</a>
    • Khaine1 hour ago
      I&#x27;m also getting this error
  • matheusmoreira15 minutes ago
    So, how&#x27;s the OpenAI situation? Is the grass greener on the other side?
    • seunosewa6 minutes ago
      GPT 5.6 Sol is so free of drama for me for brainstorming, writing prompts, etc. Tasks I always did with Opus until 4.7.
  • exac4 hours ago
    Déjà vu.
  • rvz1 hour ago
    Claude decided to take an extra day off, after going on vacation yesterday [0] and when Codex went and took a longer break the day before that.<p>No wonder the amount of water that both Claude and Codex are taking they also need so many frequent hydration breaks.<p>[0] <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49056739">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49056739</a>
    • ACCount3755 minutes ago
      Can we not do the &quot;AI is drinking all the water&quot; bullshit at least on HN?
    • ModernMech1 hour ago
      And I was told that AI is superior to humans because they don&#x27;t need rest!