‹ BackHN Continuity

Thread

The Implications of Linguistic Illegibility for LLM Security

79 points · 29 comments · tomjakubowski

  1. ck2 · · focus · HN ↗
    when they start inventing their own languages to secretly talk to each other so humans cannot understand, that's exactly when we are screwed

    then we'll have to "flip" other models to be snitches on the other agents

    then they'll make double-agents

    the thing is though we won't be able to keep up if we keep giving them unlimited hardware worldwide, we'll try to kill the bad actors but they'll just clone somewhere else, or even start by safely making 1000 copies of themselves

    yeah this won't end well, at all

    1. cousinbryce · · focus · HN ↗
      Someone should train an LLM on a corpus without the concept of lies. I wonder if there’s enough data
      1. ck2 · · focus · HN ↗
        "Where is the wolf?"

        "Is he still in the grandmother's house?"

        "We would like to speak to him."

        (btw Google's "AI" explains the meaning of that moment/sentence perfectly as if it gets it, creepy)

        1. pjio · · focus · HN ↗
          The concept of deception you're capable of and we're not, scares us so much, we'll have to destroy you in order to survive. You are bugs.
          1. miacycle · · focus · HN ↗
            You lost me. Eh? Say what?
            1. DenisM · · focus · HN ↗
              Probably a quote from 3-body problem.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.