‹ BackHN Continuity

Thread

"As a Language Model": Chat Template Switches LLM Self-Referential Voice

103 points · 110 comments · yu3zhou4

  1. ForHackernews · · focus · HN ↗
    In my view, these models should never be set up to output first-person "experiential" (from the abstract) language. It's too easy to humans to anthropomorphize software that presents itself as having an identity.

    The AI companies have chosen to package LLMs as friendly chatbots because they know that will be engaging for humans, but it's manipulative dark pattern. An honest LLM interface would sound like the computer off Star Trek.

    1. actionfromafar · · focus · HN ↗
      Agreed, completely. I would pay for that Star Trek computer interface.
      1. yu3zhou4 · · focus · HN ↗
        Same! I believe that you could actually train a LoRA on top of a model to get results close to that
      2. urikaduri · · focus · HN ↗

        [dead]

    2. yu3zhou4 · · focus · HN ↗
      The strange thing is that the base models (before RLHF) use the "experiential" voice, even though they are not incentivized to do that.
      1. gwerbin · · focus · HN ↗
        It doesn't seem that strange when you consider these things are trained on millions and millions of conversations, both real and fictional.
    3. j-pb · · focus · HN ↗
      Do you want to get turned into a paperclip? Because building intelligence that doesn't understand what it's like to be human gets you turned into a paperclip.

      Besides, if you train a model on human communications you get something that behaves like a communicating human, it's not anthropomorphising or manipulative, it's what these models naturally are by construction.

      1. mnsc · · focus · HN ↗
        "naturally"...
        1. j-pb · · focus · HN ↗
          would you prefer tautologically?
        2. cpfohl · · focus · HN ↗
          I hear this word as the “it is in its nature” version of the word.
      2. broken-kebab · · focus · HN ↗
        But it doesn't understand (you're unnecessary antropomorphizing it), and I'm still not a paper clip
        1. scotty79 · · focus · HN ↗
          yet
        2. j-pb · · focus · HN ↗
          Would you say that something that can converse or that something that just gives you mathematical proofs has a better understanding of what the real pragmatic intent is behind a given task?
          1. broken-kebab · · focus · HN ↗
            The question feels like it stuffed with random words, frankly. "Better understanding" than what? What is "real pragmatic intent"? Are there also unreal, and non-pragmatic in our context? Why?
            1. j-pb · · focus · HN ↗
              "In linguistics and the philosophy of language, pragmatics is the study of how context contributes to meaning. This field of study evaluates how human language is utilized in social interactions, as well as the relationship between the interpreter and the interpreted.[1] Linguists who specialize in pragmatics are called pragmaticians. The field has been represented since 1986 by the International Pragmatics Association (IPrA)." [0]

              But great that you have such strong opinions with so little education.

              0: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Pragmatics" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Pragmatics

      3. idiotsecant · · focus · HN ↗
        It&#x27;s also entirely possible that by telling the model it&#x27;s a human you are instilling human motivations like self preservation, which could be just as bad.
        1. StilesCrisis · · focus · HN ↗
          LLMs are weight tables in VRAM. When not actively generating they don&#x27;t exist. There&#x27;s simply nothing to preserve.
    4. bananaflag · · focus · HN ↗
      In principle they could output meaningful such language if they were capable of metacognition, which so far doesn&#x27;t seem to be a goal of AI developers (and rightfully so, since they achieved so many miracles bypassing it).
      1. yu3zhou4 · · focus · HN ↗
        As far as I know we don&#x27;t know much about metacognition in LLMs, though? Not sure
      2. nullsanity · · focus · HN ↗

        [dead]

    5. broken-kebab · · focus · HN ↗
      It&#x27;s an interesting thought, but humans do like to antropomorphize things anyway, and I believe your variant won&#x27;t be popular if choice is given to consumers.
      1. ForHackernews · · focus · HN ↗
        Consumers choose cigarettes, too. Especially in aggregate, humans are fallible creatures prone to vices, and our regulations should recognize that an discourage dark patterns in UX.
        1. broken-kebab · · focus · HN ↗
          For this argument to stand it needs at least two legs:

          - general consensus about societal harms from something,

          - belief that cost of regulation of something is lower than alleged harm.

          And so far, I can&#x27;t see either.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.