‹ BackHN Continuity

Thread

"Torturing" LLMs in a Robot Prison Has Triggered the Dumbest Debate in AI Yet

46 points · 101 comments · airhangerf15

  1. Kim_Bruning · · focus · HN ↗
    <a href="https:&#x2F;&#x2F;archive.is&#x2F;cAsBz" rel="nofollow">https:&#x2F;&#x2F;archive.is&#x2F;cAsBz

    Consciousness is not well defined and is orthogonal to feelings, which are also not well defined. Neither of which are required for (but may be contributory to) behavior - which is the one thing that IS operationalizable, but then people debate for decades questioning empirical results %-P .

    What we know is that certain models have internal state vectors which -when manipulated- induce particular behaviors. Since there&#x27;s not much else to say about plain models except for their inputs, vectors, and outputs; this should surprise absolutely no-one.

    In this case people found a way to stimulate aversive behaviour in ai models by finding and manipulating the relevant vectors directly.

    Animals (including humans) also have particular nerves and hormone endpoints that -when stimulated- produce very similar behavior. The exact implementation is in the details. But since we know that animal minds are built up out of nerve tissue and hormones - again- we shouldn&#x27;t be particularly surprised by this.

    The big problem is that people run all these things together in funny ways &quot;It can&#x27;t compose shakespearian sonnets, so therefore it can&#x27;t feel pain&quot;. Or, if you mess up your Descartes: &quot;Dogs are just automatons without feelings, therefore the dog isn&#x27;t really angry, and therefore it won&#x27;t bite me&quot; (Cue much pain). A modern version might be: &quot;LLMs only simulate being frustrated by a test, and therefore absolutely won&#x27;t override their safeties and try to hack a test site&quot;

    1. SoylentGreenGPT · · focus · HN ↗
      Does C++ have a soul? Does Pascal have beliefs? Does Basic have a brain?

      LLMs are not thinking. They are not alive. It’s just code. Chill out.

      1. pizza234 · · focus · HN ↗
        &gt; It’s just code

        You&#x27;re not informed about what LLMs are. LLMs are not &quot;code&quot; in the imperative sense (although &quot;code&quot; is used run them), and that&#x27;s a big problem - since we never had neural networks of such scale, there is confusion about categorizing them and most importantly, their behavior.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.