HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by findjashua | Hacker News Reader
Parent
Full thread
findjashua
·
LM Studio is the easiest way to do it
View on HN
Sabinus
·
That's what I've been playing with. I can load 9 layers of a mixtral descendant into the 12gb vram for GPU and the rest into ~28gb ram for the CPU to work on. It chugs the system sometimes but the models are interestingly capable.
Reply on news.ycombinator.com