~20 GB vram for the 7B model and 48 GB for the 13B model.
It depends on the context size as well. I'd recommend renting a 4090 from a cloud provider like runpod/vast ai to get started, using a PEFT tutorial.
It is mostly linear I think.
I will probably train how to fine tune on the small model but I don’t really need to use a worse model to save money.