HN
Hacker News
Top
New
Best
Ask
Show
Jobs
QLoRA 4-bit finetuning of LLMs | Hacker News Reader
QLoRA 4-bit finetuning of LLMs
github.com
7 points
·
kashifr
·
·
1 comment
Open article
Save
View on HN
kashifr
(original poster)
·
An efficient finetuning approach that reduces memory usage enough to finetune a 65B parameter model on a single 48GB GPU while preserving full 16-bit finetuning task performance!
Reply on news.ycombinator.com