‹ BackHN Continuity

Thread

Pi 1.0

1684 points · 602 comments · sergiotapia

  1. julesrms · · focus · HN ↗
    Every few weeks Pi hits #1 here and I quietly grumble &quot;mine does that too, but with lovely graphics.&quot;, so I&#x27;m saying it out loud: <a href="https:&#x2F;&#x2F;github.com&#x2F;juggler-ai&#x2F;juggler" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;juggler-ai&#x2F;juggler

    Like Pi, it&#x27;s plugins all the way down, provider-agnostic, minimal system prompts, threaded sub-agents, code-mode, multi-client remote sessions, worktrees, a context window you can actually see and edit, etc etc

    Where Pi is way ahead is the plugin ecosystem, and that takes people, which is hard to get amid the current deluge of agent action. So if there&#x27;s any spare oxygen trailing off this thread, I&#x27;d love any Pi-heads who fancy a bit of GUI action to come and kick the tyres..

    1. jonenst · · focus · HN ↗
      My personal quiet grumbling is that everyone got on the TUI bandwagon without any rational reason. I find it borderline psychotic. Or maybe humans are much more like sheep than we care to admit: we just follow the flock into the (UI) ravine. So your GUI is a nice escape.
      1. nilamo · · focus · HN ↗
        I mean, everything an agent shows to you and accepts as input is text, what does a gui even contribute to the conversation?
        1. rspeele · · focus · HN ↗
          All the agentic coding models are multimodal and can understand pictures quite well. I routinely paste in primitive paint drawings to augment my textual descriptions and show an agent what I mean, and it seems pretty effective. This can be to rough out a UI layout but it can also be useful in pure backend work drawing diagrams, or in any domain where code manipulates geometric data.

          For an example, see my MSPaint drawings in the README (scroll near the bottom) for this tiny little project: <a href="https:&#x2F;&#x2F;github.com&#x2F;rspeele&#x2F;meshorient" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;rspeele&#x2F;meshorient

          I redrew those for the README but IIRC, I drew something similar when explaining how the feature would work to the agent.

          And in reverse, after describing an architecture or an algorithm to it, I&#x27;ll sometimes ask it to draw a diagram for me to demonstrate its understanding. If it draws the picture in line with what I intended, I conclude that it has gained the necessary context to proceed. If not, I know I explained something wrong or at least insufficiently and need to provide clarification. There are some domains where a picture is worth a thousand words.

          It also helps cut through the Claudese. &quot;Your decision is needed for one edge case, found by the gate. When a T-joint meets an endcap that the mesher solves by a fan, should this bisect a fan slice or raise a warning?&quot; I&#x27;m sorry Claude, I am not from Missouri but you are going to have to Show Me this one with a picture.

          Of course you can save pictures to files and open them with external tools, but it&#x27;s nicer to have them inlined into the chat history when that&#x27;s exactly what they are: part of the conversation.

        2. julesrms · · focus · HN ↗
          It&#x27;s not just text, it&#x27;s structured text. It often contains tables, sections, lots of things that can be drawn much more nicely.

          And in juggler (and probably other harnesses too) the LLM can reply in HTML. So if I&#x27;m planning e.g. some UI changes, I might ask it &quot;show me what this button will look like&quot; and it&#x27;ll reply with a picture of that thing in its response. No temp files to open or clean up, and fewer tokens burned

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.