‹ BackHN Continuity

Thread

I built non-autoregressive decision models with RL a year ago

1363 points · 319 comments · nandakishor_ml

  1. dcow · · focus · HN ↗
    I can understand why the author feels bitter but it still feels juvenile to me. Certainly both Jev and Laya are based on the research of countless prior papers and academics. Diogo decided to build a product out of the concept. The author didn't. Publishing research papers and model weights is probably part of the problem--it feels academic. If you look at the author's profile they focus on applying AI to healthcare. Not selling general AI type safety to AI pilled companies and devs. There's a big difference there. Whether that's good or bad you can argue all day. But for the author to expect otherwise is pretty weird. I do applaud them for not stewing too much on it and trying to do something about it, though.
    1. operaopera · · focus · HN ↗
      I believe his qualms were with the "hype" in Jev's announcement: specifically calling this kind of model a breakthrough, without crediting previous art, and keeping everything closed source.
      1. prodigycorp · · focus · HN ↗
        And how is laya previous art? The project was vibecoded and posted yesterday.

        <a href="https:&#x2F;&#x2F;github.com&#x2F;NandhaKishorM&#x2F;laya&#x2F;commits&#x2F;main&#x2F;" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;NandhaKishorM&#x2F;laya&#x2F;commits&#x2F;main&#x2F;

        <a href="https:&#x2F;&#x2F;huggingface.co&#x2F;convaiinnovations&#x2F;laya&#x2F;commits&#x2F;main" rel="nofollow">https:&#x2F;&#x2F;huggingface.co&#x2F;convaiinnovations&#x2F;laya&#x2F;commits&#x2F;main

        1. nandakishor_ml · · focus · HN ↗
          The paper is one year old. <a href="https:&#x2F;&#x2F;arxiv.org&#x2F;abs&#x2F;2510.01237" rel="nofollow">https:&#x2F;&#x2F;arxiv.org&#x2F;abs&#x2F;2510.01237 <a href="https:&#x2F;&#x2F;pypi.org&#x2F;project&#x2F;hallunox&#x2F;" rel="nofollow">https:&#x2F;&#x2F;pypi.org&#x2F;project&#x2F;hallunox&#x2F;
          1. prodigycorp · · focus · HN ↗
            Excuse me, but calibrating language models to accurately reflect probabilities did not start with you.
            1. nandakishor_ml · · focus · HN ↗
              I didn&#x27;t claimed it bro it was first. Just shared the findings here.
              1. prodigycorp · · focus · HN ↗

                [dead]

                1. nandakishor_ml · · focus · HN ↗
                  just not these jev guys not opensourcing their work
                  1. verdverm · · focus · HN ↗
                    You used a number of not open source projects in your paper, are you equally upset with them for this reason?
          2. dcow · · focus · HN ↗

            [dead]

            1. jeremyjh · · focus · HN ↗
              Generally they do credit the papers and&#x2F;or people who developed the theory behind their product.
            2. charcircuit · · focus · HN ↗
              Researchers definitely do get paid for research. That Google company you mention funds research which can be integrated into their products.
              1. Matticus_Rex · · focus · HN ↗
                Sure, researchers get paid when they either do the research under contract or productize the research and sell it themselves. But they don&#x27;t usually do the latter because that&#x27;s quite difficult.
                1. verdverm · · focus · HN ↗
                  Are you grouping academic research funded by grants under the contract category?

                  To me, there is a meaningful difference and I&#x27;d add a third category, but I can also see the contract angle

            3. potterpie · · focus · HN ↗
              Ideas are cheap. But the ideas are cheap coming from anyone. The reason some ideas (like JEV&#x27;s) are taking up our attention (as opposed to Laya&#x27;s) is not because of their execution ability, but because venture funding now is subbing filling in for execution ability. I do understand your argument to this would be - &quot;welcome to the world!&quot; or &quot;that&#x27;s just how the world works&quot; - but that does not mean we do not recognize the ideas that came well before &quot;venture funding made it happen&quot;.

              I&#x27;d like to remind us all that there is a reason Joseph Liouville took the time to painstakingly review Galois’s chaotic manuscripts to credit him. It matters who did what before everyone else - if you do want to say &quot;ideas are cheap&quot; - we&#x27;d need to control for other variables before drawing conclusions.

              1. dcow · · focus · HN ↗
                I think we just have different definitions of execution. Getting VC funding and spending two years honing and SaaSifying an idea then dropping it with a hype train (yes paid for by VC funding) is certainly very different execution than someone writing a paper and publishing model weights on huggingface then vibe coding a “product” around it a year later the day after Jev is announced and the hype proves defensible and the idea is similar.
        2. prometheus1992 · · focus · HN ↗
          @prodigycorp - reading your comments here on this post - you seem pretty hurt by this post.
          1. prodigycorp · · focus · HN ↗
            Yeah, the reason why I am annoyed by it is because a person (who felt like a burner account of the laya creator) yesterday was haranguing me for saying that projects like this were vibe coded, posting the link to this project.

            I evaluated this project yesterday and found its claims un-credible. It&#x27;s literally nothing like jev. That&#x27;s some context behind why, a day later, I find it annoying that this is somehow the top story on HN.

            <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49752902">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49752902

            1. prometheus1992 · · focus · HN ↗
              What is jev like? Did they release any research paper? I really think typesafe hired someone to boost their post because there was nothing &quot;Breakthrough&quot; about their product. At least this post has some touch with the reality that this functionality was available a year ago and was well known among ML people.
              1. prodigycorp · · focus · HN ↗
                I can&#x27;t believe you say in another post that you have experience with bert and yet you don&#x27;t understand the value of a generalist classifier.

                Good models take time and effort. There wasn&#x27;t a good option for satisficers until a few days ago.

                1. jessrenoir · · focus · HN ↗
                  It is standard discourse on here if you look backwards. Attention is All You Need sounds like a big nothingburger according to this post: <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=15938082">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=15938082
              2. dcow · · focus · HN ↗
                &gt; At least this post has some touch with the reality that this functionality was available a year ago and was well known among ML people.

                When you market a product you make exciting claims relative to the audience you’re engaging with. When was the last time you saw a product marketing page reverently lost all the academic research and prior art that came together to make a product possible?

                If Layla’s functionality was available in a SaaS form in a way that could be used by all the people who are excited about and using Jev, wouldn’t this research have won hearts and minds last year when it landed? I would have a lot more empathy for the author if they’d taken a product to market and nobody cared. But even then maybe the market wasn’t ready. There are still reasonable explanations why sometimes ideas take off. We’re on a venture capital forum this shouldn’t need an explanation.

            2. verdverm · · focus · HN ↗
              I agree with your analysis based on my own last night (on another HN post to this same gripe on reddit, before this blog post). OP received a lot of echo chamber support in the subreddit, and recommended to post to HN, so here we are.

              The work is very amateurish, the &quot;paper&quot; would be a strong reject if I were still peer reviewing.

              <a href="https:&#x2F;&#x2F;www.reddit.com&#x2F;r&#x2F;LocalLLaMA&#x2F;comments&#x2F;1wijo3e&#x2F;i_literally_built_the_jev_architecture_one_year&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.reddit.com&#x2F;r&#x2F;LocalLLaMA&#x2F;comments&#x2F;1wijo3e&#x2F;i_liter...

            3. sreekanth850 · · focus · HN ↗
              I don&#x27;t really see the breakthrough in Jev. Classification, scoring, routing and returning probabilities over predefined choices are all established problems. We implemented category routing in our own retrieval system in a slightly different way: embed the incoming query, compare it against category profiles and route to the highest cosine-similarity. Obviously Jev isn&#x27;t similarity based, but the underlying task of making a constrained decision from predefined choices isn&#x27;t novel. TypeSafe says Jev has a new architecture and RLCD training, but Jev&#x27;s actual architecture, weights and training details aren&#x27;t public. So we can&#x27;t even claim Jev is specifically a BERT classifier, but also don&#x27;t see enough public technical evidence yet to call the underlying idea a breakthrough. Atleast they should publish a technical paper to prove their idea is breakthrough.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.