‹ BackHN Continuity

Thread

Vote on which of Hacker News' challenges for AI have been met

202 points · 271 comments · stabbles

  1. ben_w · · focus · HN ↗
    Very pleased one of my predictions was totally wrong: <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=23252711">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=23252711

    Sure, sure, what LLMs make still isn&#x27;t &quot;efficient bug-free code&quot;: my prediction is falsified because while LLMs can write and train new models with machine learning, ML is fundamentally not advanced enough to throw arbitraty new tasks at like this.

    1. FabCH · · focus · HN ↗
      Somewhat appropriate the site the OP links to is called „goalposts“ because as far as I can see, people keep shifting theirs.

      In your case, the comment you link to says „business tasks“ and you expanded it now to „arbitrary new tasks“. Those are not the same. An LLM today sure can do many many many business-speak conversion tasks.

      1. tripleee · · focus · HN ↗
        &gt; An LLM today sure can do many many many business-speak conversion tasks

        Not reliably, and not without supervision. That&#x27;s the main point. I&#x27;m trying really hard to figure out a workflow that doesn&#x27;t require me to review the code and I just don&#x27;t see how it&#x27;s possible (yet)

        You either need a comprehensive test suite (which requires understanding the code in order to create) or you need to review the actual implementation code to make sure it does the right thing

        1. FabCH · · focus · HN ↗
          Code is a tiny part of &quot;business&quot;.

          Most business is correspondence with people who want money from you and people you want money from.

          1. rstuart4133 · · focus · HN ↗
            The issue is that correspondence is legally enforceable [0]. LLM&#x27;s are good and getting better, but if LLMs are giving enforceable undertakings, you want to very sure they are not going to promise something that will send the company broke.

            I&#x27;m not sure what risk a businessman is willing to accept, but I&#x27;d be asking for probabilities under once in a millennium. The latest round has improved considerably in their ability to follow instructions (thank $DEITY), but they aren&#x27;t anywhere near that yet.

            [0] <a href="https:&#x2F;&#x2F;www.bbc.com&#x2F;travel&#x2F;article&#x2F;20240222-air-canada-chatbot-misinformation-what-travellers-should-know" rel="nofollow">https:&#x2F;&#x2F;www.bbc.com&#x2F;travel&#x2F;article&#x2F;20240222-air-canada-chatb...

            1. sokoloff · · focus · HN ↗
              People who demand risks to be lowered to once-a-millennium are not the type to go start or even run businesses.

              There’s nothing wrong with that, but starting a business means fading several once-a-year risks of failure and running even an established one means facing several once-a-century risks every year.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.