HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by nogajun | Hacker News Reader
Full thread
nogajun
·
Is this similar to fastllm?
https://github.com/ztxz16/fastllm
View on HN
vikmals
·
fastllm targets the GPU, while colibri uses CPU inference only
aliljet
·
I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...
Reply on news.ycombinator.com