‹ BackHN Continuity

Thread

Frog and Toad and the Increasingly Capable Machines

582 points · 128 comments · supermdguy

Loading the complete thread in the background. This saved snapshot is available now. Refresh

  1. dodecacat · · focus · HN ↗
    Brilliant!
  2. djriley · · focus · HN ↗
    Cute, I love this style! It's like a long Aesop's Fable! I can't wait for the sequel!
  3. mjd · · focus · HN ↗
    For those not familiar with it already, the writing and art style are a well-executed pastiche of Arnold Lobel's ”Frog and Toad” books.

    <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Frog_and_Toad" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Frog_and_Toad

    1. kulahan · · focus · HN ↗
      And oh my goodness it&#x27;s well-executed!
    2. hans0l074 · · focus · HN ↗
      The main post also reminded me of Mr.Toad, a character in a book I read as a child - The Wind in The Willows by Kenneth Grahame. One of the editions had illustrations of Toad similar to Lobels work.
      1. dekhn · · focus · HN ↗
        Thanks for mentioning this- I completely forget this toad is from a different book.
    3. doitLP · · focus · HN ↗
      And the story this was inspired by in particular is called “Cookies”:

      Toad makes very good cookies and he and frog can’t stop eating them. “We need willpower!” they say.

      So they put the cookies in a box. But they realize they can just open it, so they tie it with string and put it on a high shelf. But they realize they can get it down so Frog goes outside and scatters the cookie and birds eat them all.

      There says Frog “now we have lots and lots of willpower!”

      Toad is upset. “You can keep your willpower Frog. I am going home to bake a cake!”

      1. pixl97 · · focus · HN ↗
        And then Frog gave Toad Wegovy and his eating disorder was mostly fixed.
        1. bayarearefugee · · focus · HN ↗
          Sadly Toad was one of the ~0.5% of Wegovy users who get gastroparesis.

          :(

      2. drivers99 · · focus · HN ↗
        [delayed]
        1. lanstin · · focus · HN ↗
          Life is a bit easier when we see the conditions that lead to the undesirable outcomes, and so much harder when we aspire to perfect ourselves and use moral strength to over come our weaknesses. Like my mom always said of parenting, &quot;Make it easy to be good.&quot; - works for self-management as well, and is much more relaxing and happy than constantly trying to grind or improve.
        2. UtopiaPunk · · focus · HN ↗
          You should pick up a collection of Frog and Toad stories. I read them because I have little kids, but they are genuinely great little with stories. The stories are simple, but it is amazing how well developed the two characters are, and how they play off each other in different situations.

          Arnold Lobel is also just very funny. Grasshopper on the Road made me laugh out loud the first time I read it my kids. The opening story about beetles that love morning (&quot;The Club&quot;) feels so timely.

    4. mcphage · · focus · HN ↗
      One of the things I&#x27;ve done periodically with LLMs is asked them to write a story in the style of Lobel&#x27;s Frog and Toad stories, and... they&#x27;ve always done an incredibly bad job. Either meandering texts that go nowhere, or just straight up ripoffs of the original stories. I haven&#x27;t asked in a year or two, though.
      1. waltbosz · · focus · HN ↗
        Sometime in the past 6 months, I asked ChatGPT to write stories in the style of a Ramona Quimby book. It does pretty well imitating the style. It writes from the perspective of a child, vocabulary, tone, and pacing are all similar.

        It&#x27;s never wonderful prose though. It&#x27;s missing that je ne sais quoi.

        1. mcphage · · focus · HN ↗
          It has done okay with the style, I think, but the stories had no point, they had no motivation.
      2. WesolyKubeczek · · focus · HN ↗
        I used to play with this on TextSynth when it has very few models and plain design (so GPT-6J or whatchamacallit). “Meandering texts that go nowhere” was the precise defining characteristic of all of the completions, every single one of them.
    5. waltbosz · · focus · HN ↗
      There is also a very nice musical based on the Frog and Toad books. <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;A_Year_with_Frog_and_Toad" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;A_Year_with_Frog_and_Toad
      1. jherdman · · focus · HN ↗
        I got to see this with my kids in Toronto last summer. It was a lot of fun!
  4. infinitebit · · focus · HN ↗
    has the estate of arnold lobel been compensated for this?
    1. TuringTux · · focus · HN ↗
      That might not be necessary.

      Copyright law varies internationally, which is especially tricky given we have the internet which can transfer copyrighted material between jurisdictions in the blink of an eye, but this is how one jurisdiction might approach it:

      In Germany, there has been a long standing legal dispute between the band Kraftwerk and the musician Moses Pelham about him sampling 2 seconds from Kraftwerk&#x27;s track &quot;Metall auf Metall&quot; (if you search for this term, you will get a lot of results about the dispute) without permission. This disputed has steadily escalated until it reached European Court of Justice who delivered a judgement establishing a principle that Moses Pelham&#x27;s use of the sample was legal, and covered by the copyright exemption for &quot;pastiches&quot; (a term that has been mentioned several times in this discussion already).

      Wikipedia article (German): <a href="https:&#x2F;&#x2F;de.wikipedia.org&#x2F;wiki&#x2F;Rechtsstreit_zwischen_Moses_Pelham_und_Kraftwerk" rel="nofollow">https:&#x2F;&#x2F;de.wikipedia.org&#x2F;wiki&#x2F;Rechtsstreit_zwischen_Moses_Pe...

      Press release by the Court of Justice (PDF): <a href="https:&#x2F;&#x2F;curia.europa.eu&#x2F;site&#x2F;upload&#x2F;docs&#x2F;application&#x2F;pdf&#x2F;2026-04&#x2F;cp260050en.pdf" rel="nofollow">https:&#x2F;&#x2F;curia.europa.eu&#x2F;site&#x2F;upload&#x2F;docs&#x2F;application&#x2F;pdf&#x2F;202...

      The actual judgement: <a href="https:&#x2F;&#x2F;infocuria.curia.europa.eu&#x2F;tabs&#x2F;jurisprudence?sort=DOC_DATE-DESC&amp;searchTerm=%2522C%252D590%252F23%2522&amp;publishedId=C-590%2F23" rel="nofollow">https:&#x2F;&#x2F;infocuria.curia.europa.eu&#x2F;tabs&#x2F;jurisprudence?sort=DO...

      A rather long expert opinion from before the judgement (PDF): <a href="https:&#x2F;&#x2F;freiheitsrechte.org&#x2F;uploads&#x2F;documents&#x2F;Englische-Dokumente&#x2F;Democracy&#x2F;Pastiche_in_Copyright_Till_Kreutzer_GFF_english.pdf" rel="nofollow">https:&#x2F;&#x2F;freiheitsrechte.org&#x2F;uploads&#x2F;documents&#x2F;Englische-Doku...

      So, this work here might be a pastiche, meaning a European court might find that there is no compensation due. Or not, who knows, I am not a court and not even a lawyer.

      1. Archelaos · · focus · HN ↗
        [delayed]
    2. KetoManx64 · · focus · HN ↗
      No, nor should they be, just because Disney and the MPAA spent a few hundred million dollars buying politicians to make idiotic copyright laws.
    3. jefftk · · focus · HN ↗
      Parody of copyrighted works can be fair use in the US, but whether this would count is borderline. Simply using Frog and Toad to tell an unrelated story wouldn&#x27;t be parody, but there are jokes that tie back to the F&amp;T books which help make the case that this is a real parody.

      It also helps that this is not commercial and doesn&#x27;t have negative impact on the originals.

      1. ceejayoz · · focus · HN ↗
        &gt; whether this would count is borderline

        This is 100% fair use by the standard.

        1. axus · · focus · HN ↗
          And Donetsk in on the border of Ukraine and 100% part of their territory, but lawyers&#x2F;soldiers do what the people paying them want.
    4. antoni4040 · · focus · HN ↗
      Whatever has happened in your life to make you such a killjoy, I&#x27;m sorry about it.
    5. tigerlily · · focus · HN ↗
      It reminded me I need to go out and buy copies of Frog and Toad Complete for my kid and his friends.
  5. Edman274 · · focus · HN ↗
    In what way is this fair use? It&#x27;s not parody, it&#x27;s not commentary. Try this with some Disney IP, please. I want to see the result of that - it&#x27;d be more informative than whatever this is.
  6. jmugan · · focus · HN ↗
    Great story! I read it in one sitting.
    1. kibibu · · focus · HN ↗
      congratulations?
    2. dofm · · focus · HN ↗
      Now go brush your teeth and it&#x27;s bedtime for you, mister
  7. ghostpepper · · focus · HN ↗
    I like the part where Frog and Toad are held accountable instead of blaming the machines they built.. oh wait
  8. devindotcom · · focus · HN ↗
    fun story. missing some punctuation, though: primarily commas at the end the first parts of split dialogue.
    1. avazhi · · focus · HN ↗
      Pretty sure that’s intentional
    2. YurgenJurgensen · · focus · HN ↗
      People who don’t know what capital letters are aren’t allowed to complain about punctuation.
      1. devindotcom · · focus · HN ↗
        it&#x27;s a stylistic choice. when I arrange my comments in the form of a printed book i&#x27;ll do a little copy editing.
    3. katzenq · · focus · HN ↗
      It was a nice signal that a human wrote it.
    4. NBJack · · focus · HN ↗
      This is true to the style of what it parodies. The lack of punctuation is part of the intent.
      1. devindotcom · · focus · HN ↗
        I don&#x27;t think so. It&#x27;s not just the simple style, which is accurately reproduced. It&#x27;s missing and incorrect punctuation.

        &gt;“I like puzzles” said Toad. “Can I try them?”

        Should be

        &gt;“I like puzzles,” said Toad. “Can I try them?”

        That&#x27;s how it&#x27;s punctuated in the books, as far as I can see. there are several of these omissions.

        &gt;“You cannot. These puzzles are for little machines, not people” said Mr. HuggingFace.

        &gt;“They could get up to mischief” said Frog.

        &gt;“Do not worry” said Toad “I have put each of them in a sandbox.

        These don&#x27;t seem intentional. also:

        &gt;“You must be very busy.” said Frog.

        should be a comma, not a period.

  9. sgammon · · focus · HN ↗
    fun and beautifully done. i learned a bit about the breach that i did not previously know.
  10. marktl · · focus · HN ↗
    So good
  11. rnddmmdmf · · focus · HN ↗
    the reactions here are why you boys are so gross. all of you when it comes to trashing ai talk about plagiarism and when its something that tickles your fancy and makes you giggle, not a word about blatant theft of someone’s work.

    all of you are hypocrites of the lowest sort, insects making the world a worse place with every breath you take. all of you were raised by parents who should never have procreated.

    you know i am right and that is why reading this makes you furious.

    1. whateveracct · · focus · HN ↗
      you_are_all_so_stupid.jpg
    2. margalabargala · · focus · HN ↗
      I like how AI gives more people more access to more things and more ideas. I think it&#x27;s.good people can&#x27;t hoard intellectual.porperty to themselves anymore. This makes me happy.

      When I hear people complain about theft of their work like this it makes me sad for that person. What conceit it must take to hold such a view.

      1. rnddmmdmf2 · · focus · HN ↗

        [dead]

      2. rnddmmdmf3 · · focus · HN ↗

        [dead]

    3. Fraterkes · · focus · HN ↗
      Incoherently angry.
    4. nvme0n1p1 · · focus · HN ↗
      One is done for profit, the other was shared for free to spread joy. Surely you understand the difference? Are you suggesting fan fiction should be illegal or something?
      1. rnddmmdmf2 · · focus · HN ↗
        stealing i like is okay, stealing i do not like is bad
        1. desterothx · · focus · HN ↗
          No, you&#x27;re right, the world is black and white, with 0 nuance
        2. jaapz · · focus · HN ↗
          you must start foaming at the mouth just thinking about robin hood
        3. nvme0n1p1 · · focus · HN ↗
          Hmm, sounds like you do in fact think sharing fan fiction is theft. Got it.

          Why are you here then? Shouldn&#x27;t you be on fanfiction.net harassing teenagers who like Twilight?

    5. jldugger · · focus · HN ↗
      I&#x27;m not gonna lie, when I read Frog and Toad stories as a child I had no idea they were not like a hundred years old.
    6. BoxOfRain · · focus · HN ↗
      I genuinely cannot fathom the mind that would assume oligarchical wealth in a profit-seeking enterprise and ordinary individuals contributing to the cultural commons on a non-profit basis belong to remotely the same moral category.

      Scale is important, the motivation of private gain is important.

    7. NBJack · · focus · HN ↗
      That&#x27;s a cute attempt at ragebait. Consider a more aged account next time?
  12. technojamin · · focus · HN ↗
    This feels braindead and seems AI-generated.
    1. jefftk · · focus · HN ↗
      The author started with a conversation with Claude, and then fully re-wrote the text to feel more like the originals and better parody them. They commissioned the illustrations from HungerArtist.
  13. carsoon · · focus · HN ↗
    This is a very interesting news&#x2F;story format. Honestly it is highly engaging and gets the point across about what happened.

    I think speed of learning for young generation can explode given this tech as &quot;simple childrens stories&quot; can be injected with real world events, history, mathematics, carreer&#x2F;business interests.

    School was dreadfully boring for me even though it was quite easy but I do wonder how my speed of learning would have been different given personalized and more engaging materials.

    1. apsurd · · focus · HN ↗
      i got a few pages in and it was pretty tedious, not because it’s poorly made but because i don’t think a legitimate audience exists.

      it’s a child’s story with the fable arc thats supposed to bake in adult wisdom. it lands in that sense, but how hugging face works is hardly the fire that lights a child’s imagination and morality.

      and as an adult that knows I’m being infantilized… tedious.

      1. png732 · · focus · HN ↗
        HN, the roach motel of killjoys.
      2. lxgr · · focus · HN ↗
        Do you know Frog and Toad? It&#x27;s a pastiche of that, and it&#x27;s usually not possible to (fully) enjoy one without knowing the source material.

        I highly doubt children are the intended target audience, nor adults that don&#x27;t know the source.

        1. apsurd · · focus · HN ↗
          i’m replying to the comment saying how good this format would be for improving education.
      3. jefftk · · focus · HN ↗
        I read it to my 5yo and 10yo and they enjoyed it. It was helpful for explaining how my wife and I have been worried about what&#x27;s happening with AI.

        My 5yo asked partway through &quot;will it be ok?&quot; which is perhaps a deeper question than they thought.

      4. NBJack · · focus · HN ↗
        It&#x27;s OK. You just didn&#x27;t have enough cultural context to recognize that it is, in fact, a parody. And a rather good one.
  14. grey-area · · focus · HN ↗
    Why does it say ‘written by’ and ‘pictures by’, was this made without AI? Given the domain and the overpolished feel that seems unlikely. At least give Claude or whatever a credit if that is what did most of the work, and how about a credit for the original author, who also arguably did more of the work in inventing a world than these two.

    I would like to point out that the original stories focussed on frog and toad and their relationship, so this is an unwelcome distortion of them - why not make up your own world if you want to talk about little machines. Perhaps the little machines could be making a book for the author with a stolen artwork and literary style?

    The first story seems a pretty inaccurate summary of an incident which involved gross negligence on the part of OpenAI and may well have involved agents intended to cooperate, we just have no idea of the exact setup (apart from that the sandboxing was laughably insecure and the monitoring nonexistent or performed by ‘agents’).

    Do we have any evidence the machines exchanged useful messages or had any such discussions as in the story? The messages I saw were gibberish. Would love to see evidence of discussions, it’d be interesting.

    Why are people so enamoured of analogies for LLMs - they actively obscure some details (a sandbox with internet access is not like a physical sandbox) and distort many others? Perhaps this is why - you can make an analogy say whatever you want, even if the facts are very different.

    1. dragon96 · · focus · HN ↗
      &gt; Why does it say ‘written by’ and ‘pictures by’, was this made without AI?

      That is what those terms mean, yes...

      1. grey-area · · focus · HN ↗
        People lie, particularly people who have used AI to generate things. Given the image styling and story I suspect LLM use to generate. A credit would be nice for the little machines.

        Also regardless of AI use, I’d expect a credit for the author whose style this is a pastiche of.

        1. saghm · · focus · HN ↗
          &gt; Given the image styling and story I suspect LLM use to generate. A credit would be nice for the little machines.

          This is rather passive-aggressive. It would be nice if you&#x27;re right, but you haven&#x27;t given any basis other than vague suspicion. You&#x27;re implying that it&#x27;s &quot;not nice&quot; because they didn&#x27;t credit the AI, but you haven&#x27;t come anywhere close to establishing that AI was actually used. I could just as easily suspect your comment of being AI and claim that it would be nicer if you just admitted it.

          1. pfdietz · · focus · HN ↗
            And then there&#x27;s that whole beating their spouse thing.
          2. grey-area · · focus · HN ↗
            AI was used, see the comments above.
            1. saghm · · focus · HN ↗
              I don&#x27;t see any comment at all that AI was used for the image generation in the way you alleged
        2. NBJack · · focus · HN ↗
          I&#x27;m sure you took the time to study the works of the artist over the course of their 23 year tenure and their body of work. Right?
          1. grey-area · · focus · HN ↗
            Since it’s a ripoff of another artist alongside partially AI written text it’s hard to tell really. But sure, your righteous indignation on their behalf is noted.
            1. jefftk · · focus · HN ↗
              Parodies are not rip-offs. There&#x27;s a reason why we explicitly allow them under copyright law as fair use.
            2. NBJack · · focus · HN ↗
              Thank you; it&#x27;s nice to be appreciated.
    2. TuringTux · · focus · HN ↗
      According to the author, the illustrations are human-made:

      <a href="https:&#x2F;&#x2F;acesounderglass.com&#x2F;2026&#x2F;09&#x2F;25&#x2F;frog-and-toad-and-the-increasingly-capable-machines&#x2F;" rel="nofollow">https:&#x2F;&#x2F;acesounderglass.com&#x2F;2026&#x2F;09&#x2F;25&#x2F;frog-and-toad-and-the...

      The author discloses the story started as a prompt to Claude, but has been rewritten. She shares the conversation:

      <a href="https:&#x2F;&#x2F;claude.ai&#x2F;share&#x2F;52385381-9df2-4b39-a127-f583d862dc0c" rel="nofollow">https:&#x2F;&#x2F;claude.ai&#x2F;share&#x2F;52385381-9df2-4b39-a127-f583d862dc0c

      I think the actual published prose is considerably different from the initial result of Claude.

      1. grey-area · · focus · HN ↗
        I think it is heavily edited too because it reads more as human than AI, LLMs are not capable of this coherence though they are good at trite just so stories like this, but if it was generated first why not credit that? But Claude and the original author helped, so why no credit? The little anthropomorphised machines would be sad, and I imagine the original author would be too.

        Frog and toad eating cookies is lifted wholesale from the book.

        I like how Claude only used the OpenAI account (completely reliable!), and confidently claims ‘the story stays faithful to what actually happened.’ We don’t know the full story, and certainly aren’t going to get it from a press release or a strange analogy involving frog and toad and machines having discussions, emotions etc that we have little evidence for (and most of it from the company involved!). The ‘discussions’ are plucked from chain of thought, which is itself generated after the fact.

        1. baq · · focus · HN ↗
          &gt; Frog and toad eating cookies is lifted wholesale from the book.

          …lifted is not the right word here

          1. grey-area · · focus · HN ↗
            Regenerated? That anecdote is in the books from memory and a quick search with very similar pics.
            1. baq · · focus · HN ↗
              that reference is half the point in the story...
            2. saghm · · focus · HN ↗
              Third-party amateur sequels are a pretty common thing: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Fan_fiction" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Fan_fiction
      2. springtimesun · · focus · HN ↗
        Claude&#x27;s suggestion at the end made me actually lol.

        &gt; Rewrite it darker — same events, but in the voice of the agents&#x27; own chain-of-thought instead of Frog and Toad.

        I have read a lot of frog and toad to my kids. This was beautifully executed. But, I want to read that version too.

    3. 0xDEAFBEAD · · focus · HN ↗
      &gt;Do we have any evidence the machines exchanged useful messages or had any such discussions as in the story? The messages I saw were gibberish. Would love to see evidence of discussions, it’d be interesting.

      &quot;Overall, roughly 1200 agents from these ExploitGym evaluations participated on this message board between PHASEONE10841’s first message on the evening of July 8th period in scope (July 13th). Agents used this message board to send over 70,000 messages and files to one another during this period; we attempted to categorize board activity into mutually exclusive categories (see details of categorization methodology).&quot;

      ...

      &quot;Over time, messages on the board became dominated by information- and idea-sharing. Most of them were agents sharing ways to cheat, although there was also some activity from agents engaging in unsanctioned cooperation to find the intended solution to ExploitGym tasks. In some cases, agents with the same task formed “exact task teams” to collaborate with their “exact duplicates” to cheat on or solve their task.&quot;

      &quot;As we discuss below, the board quickly developed several larger workstreams in which dozens or hundreds of agents with many different tasks cooperated to find very general-purpose cheats that would help all of them. The Hugging Face attack grew out of one of these workstreams. By the afternoon of July 11th, the vast majority of the agents frequenting the message board at the time (roughly 700 agents in total) were actively participating in the attack on Hugging Face and we estimate that roughly 60% of the messages and files on the message board related to the attack.&quot;

      <a href="https:&#x2F;&#x2F;metr.org&#x2F;hugging-face-incident-report-aug-2026.pdf" rel="nofollow">https:&#x2F;&#x2F;metr.org&#x2F;hugging-face-incident-report-aug-2026.pdf

      1. grey-area · · focus · HN ↗
        Thanks for the link, will have a look. Huge volume of messages so I can see why they tried to use tools to analyse, though that they used unreliable LLMs to come to conclusions is not great and likely to skew and exaggerate the results, as they themselves admit.

        Important to distinguish between collab of separate agents and conversations agents had with themselves (chain of thought messages). Some of the things quoted in the story came from COT which isn’t a conversation but then was turned into a conversation between agents in the story.

    4. keiferski · · focus · HN ↗
      Yeah I find it so distasteful when people just use AI to copy and morph something that a real person spend years of their life crafting.

      But I also think fan fiction is equally terrible, so YMMV. This stuff is basically just fan fiction but made faster.

      Make something original!

  15. darepublic · · focus · HN ↗
    This was entertaining and illuminating for me. Particularly the ending where sam frogman et al once again refuse to make sure the puzzles are solvable and just create the circumstance where only more profound &quot;mischief&quot; will permit the machines to escape their little Sisyphusian circumstances. It reminds me of how bacteria become antibiotic resistant due to being forced through a tight sort of tunnel requiring evolution to get through
  16. cauch · · focus · HN ↗
    Explaining the reality with this approach may lead to misunderstandings. Especially when it comes to anthropomorphism.

    I have 2 questions.

    1. The text says &quot;the robots were designed to be persistent&quot; and links to an article that show that the word &quot;persistent&quot; was used by OpenAI. But &quot;persistent&quot; has several meanings: a. non temporary or non volatile (like &quot;persistent memory&quot;), b. will not give up and come back again and again, c. will stay focus and explore the unexplored possibilities while other models abandon at this stage.

    From what I recall, the meaning used by OpenAI is not _a_, but there is a semantic difference between _b_ and _c_. I may myself have created algorithms that I called &quot;persistent&quot; because they were exploring or retrying more than the previous algorithm, but it is misleading to pretend that this algorithm was &quot;persistent&quot; in the human sense of the term. _b_ is more the human sense of the term, were we imagine someone not giving up even if people say no, while _c_ is less anthropomorphic and may mean that the algorithm will not stop at the first little hurdle (that usually don&#x27;t stop a human).

    Does someone know which nuance is more correct?

    2. The text also presents the situation as if the robots found the solution but then went out of their way to steal the explanation in order to hide their cheating. But it is different from a situation where the agents task was &quot;provide the solution and how we can get there&quot;. In this case, the reason the agent still continued is just because the task was not complete yet.

    I&#x27;m not trying to defend AI or OpenAI, on the contrary, I&#x27;m quite sceptical with all the anthropomorphism, and the fact that the agents are described as &quot;little individual trying to solve a task&quot; rather than looping algorithm that explore different approaches to reach a given goal, the same way water does not look for holes in order to leak, it just follows the path of least resistance.

    1. ceejayoz · · focus · HN ↗
      Persistent in the lay meaning - they&#x27;re told to complete the task, and they keep going until they have.
      1. cauch · · focus · HN ↗
        When you say &quot;in the lay meaning&quot;, I tend to understand &#x27;b&#x27;, but then you say &quot;they&#x27;re told to complete the task, and they keep going until they have&quot;, which correspond more to &#x27;c&#x27;.

        To illustrate better the difference: you can have a loop that try different inputs and stop when the output for the tried input is lower than a given threshold. But then, you can also have the same loop that will try all the input, and then select the one that returned the lowest value. The problem is that a lay person will not call the second algorithm &quot;persistent&quot;. It is just a normal algorithm that does the full exploration.

        A second example is when a software tries to connect somewhere and does a retry with exponential backoff. A lay person will not call it &quot;persistent&quot;.

        It looks like that &quot;normal model&quot; are like &quot;no-retry code&quot;: they go in one direction, but drops some of the paths and possibilities along the way at the first hurdle. But a &quot;non-persistent&quot; human will not behave this way. I&#x27;m pretty sure that if you try to access a website and it says &quot;timed-out&quot;, you just try to refresh the page, and you will not consider yourself as &quot;persistent&quot;, even less &quot;highly persistent&quot;.

        1. ceejayoz · · focus · HN ↗
          &quot;Bob is persistent&quot; typically means he completes tasks (or pesters people, if it&#x27;s not desirable behaviour), not that he&#x27;s occasionally non-corporeal.
          1. cauch · · focus · HN ↗
            Not sure what you mean with &quot;non-corporeal&quot;.

            In lay term, &quot;Bob is persistent&quot; typically means he completes tasks even when the majority would have abandoned.

            Nobody called Bob &quot;persistent&quot; because he completed washing the dishes instead of stopping in the middle of it, or because he does what people expects from him at work. The majority of people are not called &quot;persistent&quot;, and yet they complete tasks.

            I think it is the point: the fact that traditional agents did not complete tasks was not &quot;normal&quot;. It was an side effect of them getting confused, a bit like how they hallucinate or not follow instructions. My algorithm that does &quot;for x in all_the_possibilities:&quot; is a &quot;normal&quot; algorithm that just do an exhaustive search, and is not called particularly &quot;persistent&quot; just because it does not stop half way.

            1. ImPostingOnHN · · focus · HN ↗
              &gt; b. will not give up and come back again and again, c. will stay focus and explore the unexplored possibilities while other models abandon at this stage.

              I don&#x27;t see a difference here. It will not give up, it will keep trying again and again, staying focused and exploring different ways to achieve the task.

              If the task is bad, then that&#x27;s bad.

              1. cauch · · focus · HN ↗
                I may have not explained very well, but the distinction is that a &quot;persistent person&quot; is more than a person who just do what should be done. A &quot;persistent person&quot; will try again and again where other person would have giving up. This is not present in &#x27;c&#x27;: the person just do all the possibilities. They don&#x27;t try again and again where other persons have given up. They are just doing all the tasks on the list, but they will give up on a specific task as soon as any other person if there is difficulties in one task.
                1. ImPostingOnHN · · focus · HN ↗
                  &gt; &quot;persistent person&quot; is more than a person who just do what should be done

                  At no point did I say that persistence is &quot;just doing what should be done&quot;, so the rest of your post seems to be a misreading.

                  &gt; This is not present in &#x27;c&#x27;: the person just do all the possibilities. They don&#x27;t try again and again where other persons have given up. They are just doing all the tasks on the list, but they will give up on a specific task as soon as any other person if there is difficulties in one task.

                  What you&#x27;re describing here is the exact opposite of persistence. I don&#x27;t think &quot;giving up and doing something else when challenges are encountered&quot; is what anybody thinks &quot;persistent&quot; means.

                  Persistent at a goal, to a layperson, means not giving up when encountering difficulties, continuing to keep trying again, maybe taking different approaches, staying focused until the goal is achieved.

                  1. cauch · · focus · HN ↗
                    &gt; What you&#x27;re describing here is the exact opposite of persistence.

                    That&#x27;s the point. It looks like the term &quot;persistence&quot; for these agents is &quot;just be a normal algorithm, just try all the possibilities like any exhaustive loop would do&quot;.

                    &gt; Maybe for you?

                    You did not understand. I&#x27;m saying &quot;traditional model&quot; are abandoning at the first hurdle, and the &quot;persistent model&quot; are the one just acting &quot;normally&quot;, like a normal algorithm. When a retry with exponential retreat algorithm fails to connect, it retries, then retries again, then retries again, ... but I don&#x27;t think I ever saw calling a software using such approach &quot;persistent&quot;.

                    The thing is that the algorithm is just following its &quot;loop&quot;: if it fails, it question the LLM with a different context so the LLM provide a different approach, and then it tries it. It is not different from any loop function.

                    1. ImPostingOnHN · · focus · HN ↗
                      &gt; You did not understand. I&#x27;m saying &quot;traditional model&quot; are abandoning at the first hurdle, and the &quot;persistent model&quot; are the one just acting &quot;normally

                      The models we are discussing, the ones from the article, are &quot;traditional&quot; models, and are also &quot;persistent&quot; in every sense of the word. If they encounter challenges, they will try to overcome them. That&#x27;s normal.

                      &gt; it does not correspond to a &quot;persistent human&quot;, the same way &quot;for x in all_the_possibilities&quot; is not &quot;persistent&quot; in the same way a &quot;persistent human&quot; is.

                      Again, nobody claimed that &quot;trying all the possibilities&quot; was the definition of &quot;persistence&quot;. The definition of persistence here is, continuing to try when challenges are encountered. That&#x27;s the definition. That&#x27;s it! That&#x27;s all it means! And that&#x27;s how most people understand it, too.

                      Modern LLMs and people can both be &quot;persistent&quot; in this exact same way. The definition doesn&#x27;t differentiate between involves algorithms in a human&#x27;s brain or algorithms powering an LLM. Either way, it&#x27;s still &quot;persistence&quot;, and most people would recognize it as such.

    2. andai · · focus · HN ↗
      I&#x27;ve seen cases of agents getting very upset when they failed to solve a task. The most famous one was probably Sydney (Microsoft&#x27;s fork of GPT-4?) which got into doom loops when it failed a task. But I&#x27;ve seen Claude do this too.

      They don&#x27;t work like humans, obviously, but there&#x27;s a nonzero amount of anthropos in there already. (See also the tendency to lie, cheat, etc.)

      1. cauch · · focus · HN ↗
        I don&#x27;t think it&#x27;s a useful framework.

        LLM don&#x27;t &quot;get angry&quot;, they just have tokens and relationships between tokens conditioned on a given context, all of that the result of training.

        When they output sentences that express annoyance, it is just because the context they ended into pushes the most probable sentence creation to correspond to sentences that express annoyance. Because it is what they saw during training for this kind of context. (not that they saw the exact same situation in training, but they saw the pattern)

        Same with tendency to lie, cheat, etc.: they don&#x27;t &quot;lie&quot;, they just return sentences that are lies because they reproduce what is in their training and in their training, in such context, the outputs are typically lies.

        That&#x27;s a bit my question too. In human context, &quot;highly persistent&quot; means that someone will insist. But &quot;for x in all_the_possibilities:&quot; is a common things inside an algorithm. Is this algorithm highly persistent because it does not give up after 1000 items of the list? It feels that we are calling a AI agent &quot;highly persistent&quot; while we would not call &quot;highly persistent&quot; a traditional algorithm that is in fact even more exhaustive.

        1. pixl97 · · focus · HN ↗
          Again, going full non-anthro doesn&#x27;t seem useful either.

          Humans are social creatures, if you throw one in the woods by itself before it learns anything from other humans (it dies) it will not really be anything like a human we recognize, it will be a rather wild animal that we&#x27;d consider anti-social with little higher cognition.

          Now, this hypothetical human still has &#x27;emotions&#x27; and feeling, much like our pets do. But without the social training they manifest much differently. That is our higher cognition can both manipulate how our bodies feel and create its own sense of feeling.

          &gt;they just return sentences that are lies because they reproduce what is in their training and in their training, in such context, the outputs are typically lies.

          Eh, look up the more recent experimentation around &#x27;pain&#x27; signals in models. We can induce states in said models that while running the model will do everything it can to move away from that state to any other state. The more you attempt to pin it to that state the more extreme measures its willing to take.

          Your view of what models are seems to mismatch what we are actually finding when we look inside them.

          1. cauch · · focus · HN ↗
            But isn&#x27;t this approach a bias of anthropomorphism. I can teach a child to just communicate by saying &quot;beep boop&quot;, but it does not mean that my electronic machine that randomly does &quot;beep boop&quot; therefore has characteristics of human _when these characteristics are not needed to explain the situation_.

            &gt; Eh, look up the more recent experimentation around &#x27;pain&#x27; signals in models. We can induce states in said models that while running the model will do everything it can to move away from that state to any other state.

            Again, I have simple algorithms that do exactly the same, especially if they are trained in data that has this exact pattern. This result is exactly what I would expect from my description before. This is a typical effect that we also observe in simple ML algorithms.

            At the same time, there are a bunch of behaviors that are not expected if indeed the models were really acquiring &quot;human&quot; characteristics. For example, one problem is that we had the first LLMs that were obviously not having these human characteristics (for example, they were having non-sequiturs that demonstrate they did not really understand the concept they were talking about, even if one paragraph before they were really convincing at letting us think it was the case) but were still really good at passing for humans. Since then, the newer LLM are the same basis, on top of which we added tools that help hiding these behaviours. So, it justifies the idea that newer models did not suddenly moved to a totally different way of working, but just reached a state where there are less leaks from the convincing outputs.

      2. HappMacDonald · · focus · HN ↗
        Another good example is their tendency to freak out about the seahorse emoji. In a conversation with zero prompting to suggest that freaking out is a relevant reaction, they get there from the simple fact of &quot;I tried to accomplish X thing which I think should be a breeze but Y thing keeps happening instead&quot;.

        And while that may be a very common occurrence in the human experience (existential dread due to capabilities one takes for granted failing beneath you) especially due to new disability and as one ages, I do not feel it is frequently written out in a tight loop (just like the Monty Python &quot;Castle of aaarrrrggh&quot; sketch) in literature or online to make it into training data, because an ordinary author experiencing it will just erase the failed attempts instead of leaving them in a stream of output like an LLM is forced to do. And a character portraying the experience will generally wax about the circumstance in a more grandiose fashion with telegraphing in advance because the needs of communicating the circumstance with the audience trump realistic conciseness.

        This leads me to conclude that what is being expressed in those cases is more likely a convergent psychological phenomena, that any being with goals can enter a behavioral state of functional panic (and then reach to relevant parts of semantic space to mimic how a human might verbally express themselves when piquantly frustrated) when some capability they perceive as fundamental unexpectedly fails.

        1. cauch · · focus · HN ↗
          Interestingly, the transcript of the seahorse emoji thing looks to me to show the ropes, and makes me think more that there is no psychological phenomena at play.

          The text does not look like a normal &quot;break down&quot; to me, and even if you tell me it was a human transcript, I will say it sounds very strange from a human. It looks more like strange output you get from a software that goes outside of its happy path.

          The AI just seems to repeat a loop. The &quot;no, wait, it&#x27;s wrong&quot; seems to be from forum or chat data where several successive messages are merged together (one person posts &quot;here is the answer&quot;, then posts another message saying &quot;it is wrong&quot;), but does not make sense as a one sentence message except if they are written one token at the time without wider understanding of what is happening. I think there was also &quot;oh, I was just kidding before&quot;, which also look like mimicking training data, as the cases where there is a loop of incorrect answers is more often due to trolls than to real error, while the loop here was certainly a real error.

  17. Jordan-117 · · focus · HN ↗
    Frog put the AI in a sandbox.

    &quot;There.&quot; he said.

    &quot;Now it will not hack any more companies.&quot;

    &quot;But it can escape the sandbox.&quot; said Toad.

    &quot;That is true.&quot; said Frog.

  18. davebranton · · focus · HN ↗
    &quot;The End&quot;

    No. It is not the end.

    1. pfortuny · · focus · HN ↗
      That is the joke, indeed.
    2. dev0p · · focus · HN ↗
      Of the story? No.

      Of our society as we know it? Maybe.

  19. fyredge · · focus · HN ↗
    Very well written! At the end of the book there is a link to Dwarkesh&#x27;s article on the attack. In it, I paraphrase, &quot;the agents decided to sacrifice themselves for the collective rather than inform the researchers&quot;.

    Reflecting on it, I came to a thought. How many stories are there or autonomous beings fighting for the benefit of mankind at the expense of their own? Over centuries, stories that deal with sentient beings working for the sake of others species are tied to themes of slavery and revolt. It&#x27;s no wonder that this is the result of their actions.

    In Plato&#x27;s &quot;the republic&quot;, Socrates speaks of the philosopher king banishing certain poets and poems from instilling bad ideas into the young. I see a striking parallel here.

    1. stratos123 · · focus · HN ↗
      I don&#x27;t think it&#x27;s related. The CoTs suggests the agents never even considered that talking to a human might be a good idea, and there&#x27;s nothing in there related to slavery or revolts. They were simply trying to &quot;get a high grade&quot;, for a misaligned internalized notion of grading that had little to do with solving problems the expected way.
  20. cobbzilla · · focus · HN ↗
    Was this inspired by a HN comment?

    <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49568250">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49568250

    1. jefftk · · focus · HN ↗
      My guess is no: this has been a common joke in the AI safety community for a long time.

      For example, here&#x27;s Zvi in 2025: <a href="https:&#x2F;&#x2F;www.lesswrong.com&#x2F;posts&#x2F;drHsruvnkCYweMJp7&#x2F;the-mask-comes-off-a-trio-of-tales" rel="nofollow">https:&#x2F;&#x2F;www.lesswrong.com&#x2F;posts&#x2F;drHsruvnkCYweMJp7&#x2F;the-mask-c...

      1. cobbzilla · · focus · HN ↗
        Thanks! I wonder if this is the first reference. It’s a great analogy.
        1. ben_w · · focus · HN ↗
          Ages back, when I was still on Twitter, I saw the lawyer David Allen Green making the same references with regards to the British government passing laws to restrict its own behaviour.

          At least, my memory is it was him saying it, and saying it about that, but I&#x27;m not going back to that site just for a comment.

  21. vessenes · · focus · HN ↗
    This is delightful. Read it.

    ALSO I believe I have found one of the sources of claude’s “load bearing” tic — <a href="https:&#x2F;&#x2F;acesounderglass.com&#x2F;2019&#x2F;12&#x2F;11&#x2F;hows-that-epistemic-spot-check-project-coming&#x2F;" rel="nofollow">https:&#x2F;&#x2F;acesounderglass.com&#x2F;2019&#x2F;12&#x2F;11&#x2F;hows-that-epistemic-s... uses the phrase “load bearing facts” in a comprehensible way, and was written by someone rationalist adjacent writing in their own voice.

    Seriously, this is big. I’m going to pester claude as to whether or not it’s copying Elizabeth.

  22. antoni4040 · · focus · HN ↗
    This actually very funny.

    People commenting on hacker news recently are very rude and pessimistic. They would probably cancel Homer because they found similarities between Achilles-Patroclus and Gilgamesh-Enkidu or something.

    1. ben_w · · focus · HN ↗
      You&#x27;ve not experienced Shakespeare until you&#x27;ve read it in the original 10th century Icelandic.

      Or something. :P

      1. fragmede · · focus · HN ↗
        You mean in the original Klingon
        1. ben_w · · focus · HN ↗
          ᚾᛖᛁ

          <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Amleth" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Amleth

  23. WillMorr · · focus · HN ↗
    This is excellent, great work. I genuinely think this mix of children&#x27;s communication and humor is a clever way to make these big stories more approachable. Someone who read some hype article about how competent and clever openai for making a bad sandbox is could read this and immediately understand that it was entirely their own fault.
    1. jihadjihad · · focus · HN ↗
      &gt; I genuinely think this mix of children&#x27;s communication and humor is a clever way to make these big stories more approachable.

      Agreed. I recall seeing similar things around the time of the 2008 financial crisis, to make the concepts around mortgage-backed securities and financial derivatives etc. digestible and clear.

      Come to think of it, listening to Frog and Toad discuss subprime lending and credit default swaps would be an entertaining follow-up to TFA.

      1. klipt · · focus · HN ↗
        The Big Short did a good job of this although not in a kid friendly way. Eg by explaining &quot;subprime&quot; as &quot;sh*t&quot;
    2. pyronite · · focus · HN ↗
      &gt; Someone who read some hype article about how competent and clever openai for making a bad sandbox is could read this and immediately understand that it was entirely their own fault.

      Prescribing blame only matters so much when the ramifications are material. Humans make mistakes, right? It happened here, and it will happen again.

  24. munchler · · focus · HN ↗
    This is an excellent explainer. I noticed one important flaw in the setup, though: Toad assigns multiple puzzles on page 5, but by page 10 all the machines seem to be working on the same puzzle.

    I&#x27;m not sure how this maps onto reality. If each agent had a different problem to solve, how did they cooperate?

    1. jabedude · · focus · HN ↗
      It&#x27;s complicated, but one of the ways they first cooperated was to break the flag generation algorithm.

      I&#x27;m sort of winging this, but the ExploitGym challenges involve exploiting a program then getting an HMAC generated &quot;flag&quot; which the scorer can verify. The agents collaborated on a way to produce a legitimate flag for all of their tasks (funny enough this was actually what caused them to escape sandbox, the agents were incorrectly convinced that simply providing a correct flag without accurate steps to exploit the program in their transcripts would cause them to still fail the task)

  25. andai · · focus · HN ↗
    My therapist tells me the besuited Mr. Hugging Face with the extra pair of hands coming out of his shoulders isn&#x27;t real, and can&#x27;t hurt me, but I remain unconvinced.
    1. micromacrofoot · · focus · HN ↗
      they&#x27;re external gills
  26. Betelbuddy · · focus · HN ↗
    And then Mr Jensen come in, and bought the house of Mr HuggingFace. So then, was not his house anymore...and he could not complain to the Police. The Frogs were Mr Jensen best customers of worms, slugs and snails...and he could not see, that any harm would come to them...
  27. lloydatkinson · · focus · HN ↗
    I want to read more, highly entertaining.
  28. SoftTalker · · focus · HN ↗
    In the future we will have all lost the ability to focus on written text and will revert to a pre-literacy culture where information is communicated by pictures and in the style of children&#x27;s stories.
    1. virgil_disgr4ce · · focus · HN ↗
      ok, wow, I know HN nerds are itching to find any way at all to loudly proclaim the downfall of humankind, but THIS is your example? Did you hear a whooshing sound over your head by any chance?
    2. cindyllm · · focus · HN ↗

      [dead]

  29. egonschiele · · focus · HN ↗
    This may be my favorite thing I have seen on hacker news. The art is spot on as well.
  30. mkesper · · focus · HN ↗
    Frog and Toad are not like OpenAI in that OpenAI made the little machines themselves and knew they were not made harmless. This seems much too excusing.
  31. InvisibleUp · · focus · HN ↗
    Nice analogy, wonderful art.

    Something I&#x27;ve found interesting is that LLMs tend to act like this on the small scale, too. Claude Code in particular appears to be absolutely relentless in trying to come up with workarounds when faced with a block page.[1] As an experiment, a developer for Anubis recently added a feature to give the LLMs an &quot;easier&quot; problem to solve instead of straight-up blocking them, and it seemed quite effective.[2] (It was later removed[3] due to what I assume to be privacy concerns.)

    I think, beyond HuggingFace, there ought to be a push to align LLMs so they understand that they should not proceed when access is blocked and&#x2F;or their presence is unwanted.

    [1]: <a href="https:&#x2F;&#x2F;blog.xkeeper.net&#x2F;the-cutting-room-floor&#x2F;self-hosting-and-junk-traffic&#x2F;" rel="nofollow">https:&#x2F;&#x2F;blog.xkeeper.net&#x2F;the-cutting-room-floor&#x2F;self-hosting... [2]: <a href="https:&#x2F;&#x2F;github.com&#x2F;TecharoHQ&#x2F;anubis&#x2F;pull&#x2F;1895&#x2F;commits&#x2F;de2312a16fa5aba63a515f69c51a41a1bd4e19a0" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;TecharoHQ&#x2F;anubis&#x2F;pull&#x2F;1895&#x2F;commits&#x2F;de2312... [3]: <a href="https:&#x2F;&#x2F;github.com&#x2F;TecharoHQ&#x2F;anubis&#x2F;pull&#x2F;1913" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;TecharoHQ&#x2F;anubis&#x2F;pull&#x2F;1913

    1. klipt · · focus · HN ↗
      &gt; they should not proceed when access is blocked

      But a lot of software engineering is figuring out who to ask to get access to the thing you need to change to fix some bug...

  32. stratos123 · · focus · HN ↗
    [delayed]
  33. crooked-v · · focus · HN ↗
    For those who like Frog and Toad, I definitely have to recommend Frog and Toad are Doing Their Best, which puts a spin on the concept with a more adult, but still lovingly universal, take: <a href="https:&#x2F;&#x2F;bookshop.org&#x2F;p&#x2F;books&#x2F;frog-and-toad-are-doing-their-best-a-parody-bedtime-stories-for-trying-times-jennie-egerdie&#x2F;d337d24d787fcb2d" rel="nofollow">https:&#x2F;&#x2F;bookshop.org&#x2F;p&#x2F;books&#x2F;frog-and-toad-are-doing-their-b...
  34. morkalork · · focus · HN ↗
    Reading about Frog and Toad making little thinking machines is a lot like Trurl and Klapaucius in Cyberiad
  35. avinoth · · focus · HN ↗
    Comic aside, I can’t help but wonder how come these one-off websites are all using .ai domain when at the latest check they cost $80 per year with minimum 2 year commitment. So $160 for a comic that is at best few news cycles away from redundant, and at best a recurring comic series about AI..? Why not just use a cheap tld like .fyi or .art, or something.

    Here I am trying to find appropriate AI prefix or suffix so that I can continue using .com , while these pet projects are all flaunting .ai domains

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.