Hacker News
Laya (OS Jev) on Mac M4 CoreML Offline (45 decisions per second)
EgregiousCube
|next
[-]
speedping
|next
|previous
[-]
imranq
|next
|previous
[-]
PaulRobinson
|next
|previous
[-]
LLMs that can reliably be used for control problems are the future, and I think classic/deep RL has generally been overlooked for years for a whole host of problems by wider industry because it felt inaccessible. The first thing I thought of when I saw Jev (and then Laya), was "this might move the needle in a really, really interesting way".
Local LLMs that can reliably be used for control problems smash through a lot of barriers I'm interested in, and this intrigues me a lot. Guess I'm about to become a big Laya fan if it can run on this kind of hardware to this performance.
ipsi
|root
|parent
|next
[-]
For companies? I think that's a lot more plausible, as that's mostly just a question of money - is it cheaper to run and administrate our own models, or outsource that?
For technically inclined users? I think that's unlikely unless they're able to operate on relatively cheap hardware while still being just as good as the hosted models. And by that I don't mean "a mac studio," that's far more money than I think is reasonable. A single RTX 5080, maybe, once memory prices start to drop.
frag
|root
|parent
|next
|previous
[-]
Stay tuned ;)