‹ BackHN Continuity

Thread

Fable 5 – Median thinking declined in August

428 points · 293 comments · espeed

  1. alexjplant · · focus · HN ↗
    I seem to recall Anthropic going on record saying that they don't do anything to model performance to stretch their compute capacity. I've anecdotally noticed massive peaks and troughs in performance week to week (albeit with Opus, not Fable).

    I wonder what their official explanation for this behavior is.

    1. wgd · · focus · HN ↗
      Their exact phrasing IIRC was that they "never intentionally degrade" their models.

      This still leaves an absurd amount of wiggle room for arguments like "oh no, our evals show that this quantization has no detectable effect on performance (in the eval distribution) therefore running the quant doesn't degrade quality"

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.