‹ BackHN Continuity

Thread

I built non-autoregressive decision models with RL a year ago

1363 points · 319 comments · nandakishor_ml

  1. dcow · · focus · HN ↗
    I can understand why the author feels bitter but it still feels juvenile to me. Certainly both Jev and Laya are based on the research of countless prior papers and academics. Diogo decided to build a product out of the concept. The author didn't. Publishing research papers and model weights is probably part of the problem--it feels academic. If you look at the author's profile they focus on applying AI to healthcare. Not selling general AI type safety to AI pilled companies and devs. There's a big difference there. Whether that's good or bad you can argue all day. But for the author to expect otherwise is pretty weird. I do applaud them for not stewing too much on it and trying to do something about it, though.
    1. operaopera · · focus · HN ↗
      I believe his qualms were with the "hype" in Jev's announcement: specifically calling this kind of model a breakthrough, without crediting previous art, and keeping everything closed source.
      1. prodigycorp · · focus · HN ↗
        And how is laya previous art? The project was vibecoded and posted yesterday.

        <a href="https:&#x2F;&#x2F;github.com&#x2F;NandhaKishorM&#x2F;laya&#x2F;commits&#x2F;main&#x2F;" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;NandhaKishorM&#x2F;laya&#x2F;commits&#x2F;main&#x2F;

        <a href="https:&#x2F;&#x2F;huggingface.co&#x2F;convaiinnovations&#x2F;laya&#x2F;commits&#x2F;main" rel="nofollow">https:&#x2F;&#x2F;huggingface.co&#x2F;convaiinnovations&#x2F;laya&#x2F;commits&#x2F;main

        1. prometheus1992 · · focus · HN ↗
          @prodigycorp - reading your comments here on this post - you seem pretty hurt by this post.
          1. prodigycorp · · focus · HN ↗
            Yeah, the reason why I am annoyed by it is because a person (who felt like a burner account of the laya creator) yesterday was haranguing me for saying that projects like this were vibe coded, posting the link to this project.

            I evaluated this project yesterday and found its claims un-credible. It&#x27;s literally nothing like jev. That&#x27;s some context behind why, a day later, I find it annoying that this is somehow the top story on HN.

            <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49752902">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49752902

            1. sreekanth850 · · focus · HN ↗
              I don&#x27;t really see the breakthrough in Jev. Classification, scoring, routing and returning probabilities over predefined choices are all established problems. We implemented category routing in our own retrieval system in a slightly different way: embed the incoming query, compare it against category profiles and route to the highest cosine-similarity. Obviously Jev isn&#x27;t similarity based, but the underlying task of making a constrained decision from predefined choices isn&#x27;t novel. TypeSafe says Jev has a new architecture and RLCD training, but Jev&#x27;s actual architecture, weights and training details aren&#x27;t public. So we can&#x27;t even claim Jev is specifically a BERT classifier, but also don&#x27;t see enough public technical evidence yet to call the underlying idea a breakthrough. Atleast they should publish a technical paper to prove their idea is breakthrough.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.