Train a 70b language model at home (2024)
answer.ai
answer.ai
the space is moving fast after all. they just seem to be explaining QLoRA fine tuning, (yes great achievement and all the folks involved are heroes) but reading a trending article on HN - it felt off.
turns out I was too dumb to check the date: 2024 and the title is mixing up quantized adapter fine tuning with base model training. thanks lol
Also from my experience you need more power to get some significant result. Mostly fine tuning would work if base model is very close to what you are trying to achieve and you won't be much happy with the results though.
Also context length becomes an issue trying to fit in with gpu with lesser ram.
"Please don't post shallow dismissals, especially of other people's work. A good critical comment teaches us something."