‹ BackHN Continuity

Thread

Introducing System One Models and Jev

1989 points · 520 comments · albelfio

  1. dinobones · · focus · HN ↗
    This is a good product but the naming/branding is pretty unfortunate.

    Typesafe.AI sounds like some typescript/structured output type of tool…

    What even is “system one” ?

    IMO the product/tech is really there, just needs better communication.

    1. toddmorey · · focus · HN ↗
      I mean, it's a structured output model that (apparently) can't hallucinate. I don't mind the name.
      1. flyinglizard · · focus · HN ↗
        It can't hallucinate, but it doesn't mean it can't make wrong decisions. Just because it adheres to a specific output format at all time, while LLMs have the output format at their mercy, then the claim of not hallucinating is made technically true.

        I think that this specific part is not super interesting if your harness just recovers from invalid LLM outputs.

        The latency and cost - yes, those are super interesting.

        1. mhitza · · focus · HN ↗
          You can get rigid output format from &quot;classic&quot; LLMs <a href="https:&#x2F;&#x2F;docs.vllm.ai&#x2F;en&#x2F;latest&#x2F;features&#x2F;structured_outputs&#x2F;" rel="nofollow">https:&#x2F;&#x2F;docs.vllm.ai&#x2F;en&#x2F;latest&#x2F;features&#x2F;structured_outputs&#x2F; though model support is limited.

          Would like to have something like in the original post but open weights.

    2. zenlikethat · · focus · HN ↗
      The model can&#x27;t reason comprehensively (e.g., like Sol XHigh would to solve a complicated problem), but it&#x27;s designed to be able to answer anything a human reasonably could quickly and intuitively, i.e., system one thinking: <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Thinking,_Fast_and_Slow" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Thinking,_Fast_and_Slow
      1. vintermann · · focus · HN ↗
        I wonder how well it can play chess, or go.
        1. leo4242 · · focus · HN ↗
          I was wondering the same! So I asked Astra to build me <a href="https:&#x2F;&#x2F;jev-chess-master.vercel.app&#x2F;" rel="nofollow">https:&#x2F;&#x2F;jev-chess-master.vercel.app&#x2F; where you play against Jev AI as chess player.

          You play White, and Jev plays Black. Rather than asking an LLM to generate move strings or JSON, the backend feeds all server-validated legal candidate moves into Vercel AI SDK&#x27;s experimental_evaluate(). Jev picks Black&#x27;s move and outputs its probability distribution across all legal candidates in a single forward pass (~300ms, ~$0.00004&#x2F;move).

          Github link: <a href="https:&#x2F;&#x2F;github.com&#x2F;qibinlou&#x2F;jev-chess" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;qibinlou&#x2F;jev-chess

          Give a try and let me know your thoughts! I am having lots of fun coming up with different chess strategies for Jev to try out.

          1. haute_cuisine · · focus · HN ↗
            Thanks for the demo. It plays very bad and blunders pieces on every move.
          2. ulcer · · focus · HN ↗
            I am horrible at chess and easily won. It does play very very badly.

            But it is essentially playing bullet chess right? No time to reason and is basically forced to go with it&#x27;s knee jerk reaction. So maybe it plays reasonably against another average bullet chess player?

            This is an instructive demo and makes me pause a bit when thinking about where I would trust using a classifier like this. Also there is no way to fine tune it right? So you are just left to its interpretation of the classification schema.

            Building classifiers is hard and forces you into thinking about your problem space and your comfort level with type1 or type2 errors. I fear this will encourage sloppy work because all those decisions are black boxed.

            Not like we were in a utopia careful ML applications before this.

    3. wging · · focus · HN ↗
      I had a different initial confusion - it seems this company has no relation to the company formerly known as Typesafe <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Akka.io" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Akka.io
    4. salicideblock · · focus · HN ↗
      &gt; What even is &quot;system one&quot;

      I definitely agree it&#x27;s underexplained in type safe.ai&#x27;s materials.

      I have to assume it&#x27;s a reference to the fast, heuristic, intuitive &quot;system 1&quot; process in humans, as opposed to the slow, procedural, reasoning &quot;system 2&quot;.

      This theory is recognized, among others, in Daniel Kahneman 2002 Nobel prize on Economics.

      1. ostacke · · focus · HN ↗
        You are correct, they state that on their website.
    5. ArafatMu · · focus · HN ↗
      Struggled to understand the use case till after i came across this <a href="https:&#x2F;&#x2F;eu.36kr.com&#x2F;en&#x2F;p&#x2F;3991165197188099" rel="nofollow">https:&#x2F;&#x2F;eu.36kr.com&#x2F;en&#x2F;p&#x2F;3991165197188099 article, which referred to them as the &quot;referee for agents.&quot;
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.