Tuning and Testing Llama 2, Flan-T5, and GPT-J with LoRA, Sematic, and Gradio
sematic.dev
sematic.dev
You can use this https://github.com/PygmalionAI/training-code
Or, you can use this; for QLoRA https://github.com/artidoro/qlora
The tools and mechanisms to get a model to do what you want is ever so changing, ever so quickly. Build and understand a notebook yourself, and reduce dependencies. You will need to switch them.
did you know that the weight adapter happens on demand, on inference, in the LoRA forward pass function?
leaky, leaky, leaky
GGML will save us, surely
Their unwavering commitment to open-source should be celebrated by all tech enthusiasts. Not sure why people poo-poo on them.
At some point, all the nice things they offer for free or cheaper will go away or become expensive.
Thank you Huggingface!
feed it a list of YouTube/SoundCloud quality “artists + song titles” and ask it to clean them up/figure out how split/parse them into CSV or JSON and then identify their genre
I want to make sure I’m not being too harsh when I criticize these as useless if they can’t do this “basic” task because I’m pretty sure I was able to get GPT-3.5 to do this reasonably well for about $0.50 with no token cost optimization
I’m just curious why people are so infatuated and putting so much effort into all of these other open source models if they couldn’t complete this basic task.
I’m curious what could be done time investment wise by a single person like myself that could “tweak” LLAMA2 into being able to do a task it by default can’t
GPT4 over the API is too fine tuned, which constrains its behavior. It fails to capture nuance in instructions. When you have the bag of weights, you can actually control your model. Having actual control over the model, and understanding the infrastructure that it's running on helps you meet actual SLAs.
And it's cheaper, if you're not backed by infinite venture money.
I don't see how this warrants the extra exciting popularity of LLAMA2, etc.
I still haven't found my own personal niche "good enough" test case
- Some people are committed to open source.
- Some people want to play with/learn/modify the technology, not just use it instrumentally.
- Some people want to play with these models without the surveillance that comes with renting them on OPC.
I'm sure there are other reasons I'm not thinking of, but I'm in the middle of that particular Venn diagram.