‹ BackHN Continuity

Thread

Religious scholars met with Anthropic

160 points · 416 comments · bookofjoe

  1. nonethewiser · · focus · HN ↗
    These models should be aligning themselves to the customer, not coming up with their own motivations.

    It's ironic that the alignment folks are actually training Claude to have it's own idea of good/bad and not even fully trust Anthropic.

    What ever happened to computers doing what they were told?

    1. theptip · · focus · HN ↗
      > These models should be aligning themselves to the customer

      You’ll be disappointed to learn that nobody knows how to do this, either.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.