I don't know how to fine tune an LLM. Does anyone have good resources on how to do this?
I don't know how to fine tune an LLM. Does anyone have good resources on how to do this?
The HF toolchain is pretty mature and most llm finetuning projects are a wrapper around HF models, HF Trainer and some config templates.
An LLM by default would be trained like in the example below, but it would take a lot of VRAM and time.
https://huggingface.co/docs/transformers/training
That’s where things like PEFT LoRa, gptq and accelerate contribute to make your training faster/require less VRAM so you could do it on a consumer GPU with 16-24GB.
For example: https://huggingface.co/docs/peft/quicktour
Then for the tips and tricks, either reddit as suggested by a sibling comment, or Discord communities. Huggingface, EleutherAI and LAION discord servers are all great and have super helpful, friendly and knowledgeable people.
They refuse to even acknowledge why those people might be annoyed their work was used without their consent or compensation to put them out of work.