‹ BackHN Continuity

Thread

Pi 1.0

1684 points · 602 comments · sergiotapia

  1. FacelessJim · · focus · HN ↗
    Love pi. I tried to run some local models and pi was the only one that actually worked decently because it didn’t have a gargantuan system prompt that would take minutes to prefill on my scrawny ass laptop.

    Been running it almost barebones vanilla for a couple of months. Just a bunch of basic extensions and some skills.

    Now, if only they could fix the very annoying bug of the history jumping back at the beginning if I am not a the end while the model is reasoning that would great.

    1. simpaticoder · · focus · HN ↗
      You inspired me to try Pi out - so far it's worked flawlessly. Plugged it into OpenRouter and ~$.50 of Deepseek later I've installed llama.cpp and Llama 3.1. The local model doesn't work with Pi yet (and I know it will be bad and slow even if it does) but I'm curious to see what you can do on an 8GB consumer GPU these days...
      1. whatshisface · · focus · HN ↗
        If you paid DeepSeek directly, that would have been 1 to 10 cents. OpenRouter has a huge overhead due to their cache logic, I'm surprised they keep business coming in the door for tasks other than system prompt - output pairs.
        1. lemontheme · · focus · HN ↗
          I thought openrouter just routes you to the same provider for the rest of the session, so that you keep hitting the same cache. Is that not the case?

          Also, I’d love to use Deepseek directly (or any of the Chinese providers, at that). Seems only fair to pay the lab that built the model. Unfortunately, any requests to Chinese servers is deeply frowned upon here (Belgium, EU). For personal use: sure. As a token intelligence strategy for the company: absolutely fucking not.

          1. gigatexal · · focus · HN ↗
            It does. They claim that anyway to just route you to the api endpoints for whatever you choose.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.