Ollaya – Ollama for open-source, Jev-style decision models
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Ollaya – Ollama for open-source, Jev-style decision models
Unofficial Hacker News client; not affiliated with Y Combinator.
solaire_oa · · focus · HN ↗
Like, their example is of classification for a support interface.... `refund_requested`. Pretty convenient bool given the example is about a refund- what if 99% of submissions don't ask about a refund? Also, is that user not a `churn_risk`? What could possibly qualify as a churn risk if not a user asking for a refund?
<a href="https://ollaya.dev/library/laya" rel="nofollow">https://ollaya.dev/library/laya The examples suffer the same problem of why I'd prefer to use a string column vs an enum. Changing an enum means you need to update the db, using a string you can do whatever.
I'm not trying to be negative, I genuinely want to know about some practical examples (that don't require tons of backwards maintenance).
devttyeu · · focus · HN ↗
Really I think "smart grep" is a pretty good one ('look for an error looking vaguely like this'). Also I think sql-based shell history + decision model is quite good to make the last 'which one of those choices is best fit given users past few commands' etc.
spaniard89277 · · focus · HN ↗
devttyeu · · focus · HN ↗
But with Jev you're just paying for input (prefill) which is really fast, and in case of Jev specifically costs 50% of Deepseek V4.1 Flash (which has famously really cheap input token pricing).
I put 250MB / 1M lines of logs through Grev and it cost ~$10USD, DSv4.1 would be at least 10x that and much, much, much slower. With Jev/Grev that 1M requests took 10 mins
Edit: completely misread your question - yeah you could finetune specialized models to do that, probably based on some decent pretrained llm base, that is true for roughly any Jev-shaped problem. Do you want to bother doing that, also having to deal with having to host a zoo of specialized models?
motoboi · · focus · HN ↗
solaire_oa · · focus · HN ↗
I very much appreciate your to-the-point, non-vibed README as well, ty for that.
devttyeu · · focus · HN ↗
On the readme I'm so sorry to tell you that, but it's 100% written by Opus 5.5 with zero "pretty please don't write slop" prompting, it's just how slop is going to look like from now on. I've been writing code for 15 years or sth like that and the code is also what I'd call pretty reasonable..
solaire_oa · · focus · HN ↗
solaire_oa · · focus · HN ↗
But if that were solved, I could see giving ollaya/grev to LLMs themselves, giving LLMs their own massive token-saver.
cobanov · · focus · HN ↗
[dead]
motoboi · · focus · HN ↗
And if you have not been, it’s for when you have to extract the context from text. When you have numbers or fixed options, it’s just a matter of code.
So if you find yourself having to decide if a given user comment is a refund_request, that’s for that.
It’s not perfect, you still have to fine-tune (or calibrate) using examples you have (and keep those examples updated over time). But it’s way better than trying to parse text with regexes.
[deleted] · · focus · HN ↗
[deleted]
[deleted] · · focus · HN ↗
[deleted]
colordrops · · focus · HN ↗
maskedpirate · · focus · HN ↗