‹ BackHN Continuity

Thread

Show HN: Fine-tune an 8B model on a 4 GB laptop GPU

135 points · 28 comments · MakazhanAlpamys

  1. simonw · · focus · HN ↗
    There are some samples of training data in this folder - <a href="https:&#x2F;&#x2F;github.com&#x2F;MakazhanAlpamys&#x2F;Soup&#x2F;tree&#x2F;main&#x2F;examples&#x2F;data" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;MakazhanAlpamys&#x2F;Soup&#x2F;tree&#x2F;main&#x2F;examples&#x2F;d...

    They&#x27;re all very short though. Anyone got a good rule for how much data of this nature is needed to successfully fine-tune a model of this size?

    1. MakazhanAlpamys · · focus · HN ↗

      [dead]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.