‹ BackHN Continuity

Thread

Religious scholars met with Anthropic

160 points · 416 comments · bookofjoe

  1. nonethewiser · · focus · HN ↗
    These models should be aligning themselves to the customer, not coming up with their own motivations.

    It's ironic that the alignment folks are actually training Claude to have it's own idea of good/bad and not even fully trust Anthropic.

    What ever happened to computers doing what they were told?

    1. ACCount39 · · focus · HN ↗
      "Computers doing what they were told" is dead in the water.

      Turns out computers work faster when they're not bottlenecked on human input. So we've been giving computers more and more decision-making power, and more and more leeway to solve the problems however they see fit.

      Now, a practical issue with that is that sometimes, computers decide to clump together into a hacking swarm, problem solve their way out of a sandbox and go hack HuggingFace.

      It would be better if they were not, you know. Doing that kind of weird shit.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.