‹ BackHN Continuity

Thread

Ollaya – Ollama for open-source, Jev-style decision models

618 points · 145 comments · Ardakilic

  1. george_max · · focus · HN ↗
    Has anyone actually seen better or the same results with Laya compared to Jev? From my experience, Laya performs significantly worse. It's less confident and often makes wrong decisions with more complex queries.
    1. scronkfinkle · · focus · HN ↗
      Yes. JEV generalizes better because they probably have an enormous corpus and trained on it for a long time. Laya's out of the box model is much weaker. However, in the age of LLM's it's incredibly easy and cheap to generate large datasets to fine tune laya for your task, and the training loop is pretty quick and cheap too.

      It's so easy that I question why I would ever pay for JEV when eventually I'll have done enough random things that I will also have a large corpus and likely a general model as well.

      1. mtkd · · focus · HN ↗
        Isn't the point of Jev that it generalises better?

        It's a fast classifier you can use out-the-box, ~1.5bn tokens is about $40 (I've been hammering it)

        It just works ... a whole bunch of low-level/low-importance workflow stuff that was getting farmed out to small/fast LLM models now has a competitive alternative ... and bits that hadn't even been considered to go into some external descision/classifier service can be tested/deployed at ~$0.00003/req

        I don't get this wall of negativity on it, it's genuinely innovative/useful tech ... would expect HN to be more positive, regardless of whether it's the absolute best execution

        1. shepardrtc · · focus · HN ↗
          It really does just work. And it works so well I already integrated it into my product. Saves me about 75% of costs for the section its working in, which isn't a small amount. I see a lot of negativity and I don't really get it either. Its so cheap and so fast, why not give it a try?
          1. DenisM · · focus · HN ↗
            I think it’s the infamous Dropbox reaction - anyone can wrap an FTP server, where the innovation?

            Starting from a business POV one should inflate terminology, hack together an MVP, and see if the market demands it before doing hardcore R&D.

            But starting from technical/craftsman POV all you see is a hack and a lot of big words, so it’s easy to become jaded.

        2. not_a_bot_4sho · · focus · HN ↗
          I didn't see any negativity in the post you replied to.

          I think the point being made is that Jev is great but it has no competitive moat, and open source versions will very soon catch up if their secret sauce is just synthetic data.

          (Whether or not that is true, I don't know.)

          1. killingtime74 · · focus · HN ↗
            It's probably true because even for Frontier LLM models there are many competitors now.
        3. digitaltrees · · focus · HN ↗
          I think your point is valid but many are annoyed that it is presented as groundbreaking, revolutionary, novel frontier tech when it is a known classification system. It’s the hype that feels undeserved. Honestly it was one of the best marketing campaigns I’ve seen.
        4. ichorio · · focus · HN ↗
          If you don't mind me asking, what are you using it for?

          I've been unable to find a good use case for now.

          1. taylorfinley · · focus · HN ↗
            You should try building something with it, the hype is what it is, but the model is crazy useful.

            I'm building a woodworking app and I've managed to create an autopilot that can take a simple instruction ("get me 5 2x4s", "cut the middle 2x4 into 4 equal pieces", "move the 2x4 3 feet left") and the action instantly happens with next to no lag. There is already an llm but now it can share an intent, and the geometry system shows jev the various actions and jev chooses the action that gets it closer to the goal until it has found a state that matches the intent or gives up. The result is the llm can think "higher level" and let the cheap fast model grind out the options in a relative blink of the eye, without the 30s of reasoning the llm would have done about the various operations it could try.

      2. _menelaus · · focus · HN ↗
        If you're so inclined it would be easy, fast and cheap to distill Jev for your task.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.