Laya on Mac M4 CoreML Offline
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
Laya on Mac M4 CoreML Offline
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
PaulRobinson · · focus · HN ↗
LLMs that can reliably be used for control problems are the future, and I think classic/deep RL has generally been overlooked for years for a whole host of problems by wider industry because it felt inaccessible. The first thing I thought of when I saw Jev (and then Laya), was "this might move the needle in a really, really interesting way".
Local LLMs that can reliably be used for control problems smash through a lot of barriers I'm interested in, and this intrigues me a lot. Guess I'm about to become a big Laya fan if it can run on this kind of hardware to this performance.
frag · · focus · HN ↗
Stay tuned ;)
putna · · focus · HN ↗
EagnaIonat · · focus · HN ↗
To me it’s like a solution looking for a problem that is already solved.
bigyabai · · focus · HN ↗
If BeRT had any potential to disrupt the datacenter buildout, it already would have.
viraptor · · focus · HN ↗
ipsi · · focus · HN ↗
For companies? I think that's a lot more plausible, as that's mostly just a question of money - is it cheaper to run and administrate our own models, or outsource that?
For technically inclined users? I think that's unlikely unless they're able to operate on relatively cheap hardware while still being just as good as the hosted models. And by that I don't mean "a mac studio," that's far more money than I think is reasonable. A single RTX 5080, maybe, once memory prices start to drop.
Izmaki · · focus · HN ↗
bigyabai · · focus · HN ↗
So, extrapolating from your gaming example, it will take smartphones only... *checks clipboard* ...100 years to achieve datacenter-level performance at the pace of 2010's improvements.
mynegation · · focus · HN ↗
bigyabai · · focus · HN ↗
Izmaki · · focus · HN ↗
Your argument is true if the number of parameters in an LLM is the only measurement for quality - a bit like number of bolts per aircraft or lines of code in software. I'd bet you that an 8b parameter LLM will in 5 years outperform the datacenter-level LLMs of today.
itemize123 · · focus · HN ↗
oezi · · focus · HN ↗
[dead]
tentacleuno · · focus · HN ↗
ryuuseijin · · focus · HN ↗
[1]: <a href="https://github.com/mizorewww/laya-coreml" rel="nofollow">https://github.com/mizorewww/laya-coreml
putna · · focus · HN ↗
frag · · focus · HN ↗
putna · · focus · HN ↗
aidiveyt · · focus · HN ↗
[dead]
altano · · focus · HN ↗
WASDx · · focus · HN ↗
putna · · focus · HN ↗
ImJasonH · · focus · HN ↗
imranq · · focus · HN ↗
putna · · focus · HN ↗
jwpapi · · focus · HN ↗
adinb · · focus · HN ↗
jwpapi · · focus · HN ↗
adinb · · focus · HN ↗
EgregiousCube · · focus · HN ↗
putna · · focus · HN ↗
physicallyIllfr · · focus · HN ↗
speedping · · focus · HN ↗
brcmthrowaway · · focus · HN ↗
putna · · focus · HN ↗
grejioh · · focus · HN ↗
[dead]
tc3oliver · · focus · HN ↗
[dead]