‹ BackHN Continuity

Thread

I quit OpenAI because its culture is broken

473 points · 791 comments · Brajeshwar

  1. flatline · · focus · HN ↗
    There are numerous problems with “alignment.” What are “human values” to begin with? He outlines some at the beginning of the post, implicitly: build bigger, better, more powerful things faster without adequate safeguards. We are literally pouring trillions of dollars of value into this enterprise, and I would say this is something that many humans also value in a qualitative sense. Then we have explicit values which in the West are largely rooted in Christian morality. Nietzsche circled this dichotomy two hundred years ago and I feel like what we have gotten since then is an increasingly detailed anatomy of power as the basis for what is normal vs deviant behavior. He who has the power, makes the rules, to be reductive.

    I do think this carries some weight from this particular author due to the length of his tenure. I happen to agree with him in spirit, but this is still largely a post revolving around sentiment not substance. Does anyone think that the overriding incentives even leave room for something like this in practice?

    1. none_to_remain · · focus · HN ↗
      Glaringly elided problem of "aligned with who?" when the user, the model creator, the government, and various other parties can all be lined up different ways. If I want the recipe for meth and the robot won't tell me, that's misalignment from my perspective.

      At least the Rationalists will handwave something for that with their "coherent extrapolated volition" idea where the superintelligence is supposed to figure out what humanity would collectively want if humanity was superintelligent and good, not that I buy it. This guy seems [.] to be coming from the NGO blob world.

      [.] <a href="https:&#x2F;&#x2F;david.robinsonian.com&#x2F;assets&#x2F;pdf&#x2F;dgr_cv.pdf" rel="nofollow">https:&#x2F;&#x2F;david.robinsonian.com&#x2F;assets&#x2F;pdf&#x2F;dgr_cv.pdf

      1. ben_w · · focus · HN ↗
        I don&#x27;t buy CEV either, but the Rationalist answer on this topic is that while CEV stops some future super-AI literally killing everyone because a user forgot to specify one minor clause that they thought was obvious in a mundane wish…

        … nobody knows how to actually make an AI would do CEV.

        1. MichaelZuo · · focus · HN ↗
          Since there can be no universal arbiter on “value”… the whole concept of “CEV” is nonsensical to begin with.

          It’s a big pretend game.

          1. ben_w · · focus · HN ↗
            Mm.

            I agree it suffers from the same problem as any other form of utilitarianism (i.e. what even is the utility function). Or indeed all forms of ethics, because for basically every topic there&#x27;s at least two cultures which disagrees with each other.

            On the other hand, even as a toy model (something I can also say for all ethics), CEV seems like it might be a step in perhaps a useful direction: &quot;When a user asks you do do something, first figure out what they actually meant to ask you if they were smarter, then do that instead&quot; is better for the user than just &quot;do the thing&quot;, though it still has problems with &quot;what happens when the thing they want is illegal?&quot;

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.