‹ BackHN Continuity

Thread

I built non-autoregressive decision models with RL a year ago

1363 points · 319 comments · nandakishor_ml

  1. cjalmeida · · focus · HN ↗
    >Zero-shot vs. Fine-tuning: Out-of-the-box base models score ~0.35 on the typed-decisions benchmark (near random). The 0.766 score is achieved by fine-tuning on the benchmark's train split. Treat Laya as a fast foundation model to specialize, not as an omniscient zero-shot oracle.

    This should be way up in the article. Fine tuning is a pain, requiring it for good results put Laya in a whole different category vs Jev

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.