‹ BackHN Continuity

Thread

Grok 4.7

609 points · 541 comments · meetpateltech

  1. Tsarp · · focus · HN ↗
    Waiting on simonw "Generate an SVG of a pelican riding a bicycle " benchmark to judge this model
    1. rvz · · focus · HN ↗

      [dead]

      1. user43928 · · focus · HN ↗
        You don't think it's useful to learn whether a model's "intelligence" generalizes beyond the tasks and modalities it is usually optimized for?
        1. TylerE · · focus · HN ↗
          Absolutely not. Makes about as much sense as judging a car based on how good an airplane it makes.
          1. user43928 · · focus · HN ↗
            I disagree. If GPT-7 can draw the Mona Lisa in MS Paint via computer use, this would be interesting.

            That it isn't the most efficient way to achieve the same end result is irrelevant.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.