‹ BackHN Continuity

Thread

Livenerf: Has Opus 5.5 been nerfed yet?

922 points · 392 comments · bryan0

  1. msejas · · focus · HN ↗
    As a claude code power user, when I get the 'rate the feedback on Claude' pop up, I used to say good or fine out of habit, and immediately after sending this feedback, I felt an instant degradation and mistakes that usually don't happen.

    Now I dismiss it every time and the quality is more consistent.

    Complete adhoc and personal experience but something I've observed, wouldn't be surprised if they nerfed on a per session basis

    1. itopaloglu83 · · focus · HN ↗
      Another personal anecdote, but I had the case where Claude would visibly improve after a negative feedback.

      And strangely, expressing frustration multiple times in a row would reliably trigger a feedback popup as well.

      1. okwhateverdude · · focus · HN ↗
        When the source maps leaked for the harness, it was revealed that they track how often tool use is rejected and how many times you say "fuck". They definitely try to track frustration sentiment.

        That said, given the propensity for mature code bases to have &quot;fuck&quot; in commit messages&#x2F;comments and those are typically of higher quality, I curse up a storm when the clankers make mistakes, if only to try put more quality-code valence into context. <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=36584464">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=36584464

        1. ClikeX · · focus · HN ↗
          To be fair, I also do my best work when I get to swear like a sailor.
        2. wunderlotus · · focus · HN ↗
          True, but tracking ≠ Claude using that feedback to improve performance right there and then
        3. jcutrell · · focus · HN ↗
          As a hiring manager I now have a rational reason to prefer people willing to use profanity. Thank you.
    2. Galilyou · · focus · HN ↗
      Wow .. really! Definitely testing that out a few times meself
    3. xnorswap · · focus · HN ↗
      This is &quot;rub your gameboy the right way to catch more pokemon&quot; levels of insanity.

      Random performance is random, your brain will jump through hurdles to fit patterns where there aren&#x27;t any.

      1. unglaublich · · focus · HN ↗
        On the one hand, yes, on the other hand it would also not be too hard to do routing based on such a parameter and Antrophic has shown (through Fable and Mythos) that they have such a transparent routing system in place.

        In fact, since they have some rule based system (if bioenegineering or security, route to degraded model) it would be almost trivial to add &#x27;user has filed feedback&#x27; to it.

        Not saying this is what happens, but just that it&#x27;s not as insane as it sounds.

        1. ftchd · · focus · HN ↗
          Not unglaublich
        2. smashed · · focus · HN ↗
          Yup.

          And it&#x27;s also the kind of solution a misaligned agent would implement:

          Make sure users are happy about their experience but also optimize resources.

          The obvious solution is to route users based on feedback.

      2. bb123 · · focus · HN ↗
        I could totally see them running an experiment to test user&#x27;s stickiness&#x2F;quality perception with decreased model performance. Facebook was doing exactly this in 2016. They tested the loyalty of Android users by secretly introducing errors that would crash the app to find the threshold at which a person would abandon. That was 10 years ago - imagine what the state of the art in user manipulation is like now.
      3. jayd16 · · focus · HN ↗
        Its truly naive to think this is impossible or even improbable. You can say there&#x27;s no proof or that its not the case but certainly much more has been done to juice 5 star ratings.
    4. Aurornis · · focus · HN ↗
      As a counterexample, I click that feedback all the time and nothing ever changes afterward.
      1. SilverSlash · · focus · HN ↗
        My personal issue with clicking the feedback is that I&#x27;m doing free RLHF for Anthropic without getting anything in return.
        1. Aurornis · · focus · HN ↗
          You’re probably participating in 100 or more A&#x2F;B tests per year without knowing it and getting nothing in return other than a minor contribution to the company’s internal knowledge about what customers like.
          1. redanddead · · focus · HN ↗
            That sucks
          2. grim_io · · focus · HN ↗
            There is a difference.

            Using feedback on your sessions makes that session, and presumably all attached data, fair game for training.

            1. margalabargala · · focus · HN ↗
              As if it wasn&#x27;t already? The feedback button may tag it a certain way but the AI company certainly isn&#x27;t going to let that sweet sweet user interaction go to waste just because they didn&#x27;t rate it.
              1. grim_io · · focus · HN ↗
                If contracts didn&#x27;t matter, sure I guess.
                1. margalabargala · · focus · HN ↗
                  Are you under the mistaken impression that hitting a 1, 2, or 3 for bad&#x2F;fine&#x2F;good is itself sufficient permission to allow Anthropic to train off your chat?

                  They explicitly ask, after the rating, whether they can look at your chat. You can just say &quot;no&quot; to that, if you believe that they are following the rules they say they are.

                  Your original statement of &quot;using feedback gives them permission to train&quot; is just plain false.

                  1. grim_io · · focus · HN ↗
                    Truth be told, I never used the feedback function because of the clauses that reference the functionality in connection with training&#x2F;service improvement.
        2. iamjackg · · focus · HN ↗
          How is it RLHF if they don&#x27;t get a transcript of your session?
        3. Aachen · · focus · HN ↗
          A better product isn&#x27;t something?
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.