‹ BackHN Continuity

Thread

I built non-autoregressive decision models with RL a year ago

1363 points · 319 comments · nandakishor_ml

  1. iamflimflam1 · · focus · HN ↗
    Probably important to call out this part of the post:

    Zero-shot vs. Fine-tuning: Out-of-the-box base models score ~0.35 on the typed-decisions benchmark (near random). The 0.766 score is achieved by fine-tuning on the benchmark's train split. Treat Laya as a fast foundation model to specialize, not as an omniscient zero-shot oracle.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.