5 comments

  • trashymctrash49 minutes ago
    Opus 5.5 has reduced this to an acceptable level for me. Are you also using this plugin with that model?
  • citizenfishy40 minutes ago
    Claude Mods seem to answer this pain better without affecting the model reasoning
  • TZubiri44 minutes ago
    Read up on Chain of Thought<p>The model is essentially thinking out loud, when you ask it to be more concise, you make it think less, therefore producing more erroneous answers.<p>Some models have an internal chain of thought (claude being one of them), which sometimes isn&#x27;t even published to avoid reverse engineering, but it seems that this might still be a problem.<p>What you&#x27;d want actually is a layer that summarizes the actual answer, but that&#x27;s actually an internal prompt by claude that you are not seeing, the model just doesn&#x27;t expose the necessary bits for you to hack this together.<p>Try another model that exposes the raw llm output instead of exposing a CoT result directly.<p>Of course the real hack is learning to read diagonally without reading every single word, this is a skill that is useful in general. It&#x27;s also less effort in general, instead of making plugins and super customizing the thing, you just consume the default settings, which are hyperoptimized, and require no time spent in configuration.
    • cyanydeez38 minutes ago
      I think you&#x27;re humanizing too much.<p>The &lt;think&gt; blocks are an attempt to explore the gradient descent space to escape local minimums and find a better global minimum to continue the descent.<p>While verbosity _might_ do this better, you could easily consider things like &quot;but wait am I forgetting ....&quot; as just one token. So if you actually do it right, you could replace all that with a &quot;hold on&quot; or something of a terse variety.
  • aidiveyt5 minutes ago
    [flagged]