S-LoRA: Serving Concurrent LoRA Adapters
github.com
github.com
https://en.wikipedia.org/wiki/LoRa
Instead it's about LoRA, note the capitalized last A, or Low-Rank Adaptation, a method for tuning LLMs.
For further reason to click on the link, it's a consumer grade long range spread spectrum wireless communication technology that's been gaining prominence in recent years.
Multiple lora adapters? Different bands, different channels, spreading factors? Like an ultimate meshtastic relay covering both 433 and 868mhz, but with some simpler logic?
Nope... disappointment.
‘Conventional’ (if that means anything in a field 10 minutes old) wisdom is “fine tune to add knowledge, LoRA to adjust presentation” - could you comment on your experiences with this?
Both tuning and LoRA is more about how it responds rather than what it knows
https://github.com/predibase/lorax
Does a similar thing and seems more active.
I plan on releasing my next iteration of my model (trained on generating knowledge graphs, among other things) as both a lora, and a pre-merged model for multiple base models (Mistral 7b, Llama2 13b, Phi2 2.7b, Yi 34b, and possibly Mixtral)
Will be interesting to compare the results between them, considering last time I just did Llama2 13b. Can't wait to see the improvements in the base models since then.
e.g., https://huggingface.co/tloen/alpaca-lora-7b https://huggingface.co/winddude/wizardLM-LlaMA-LoRA-7B https://huggingface.co/MBZUAI/bactrian-x-llama-7b-lora
But HF is a great place to download stuff, but doesn't really offer much for discoverability. How you find LoRAs you want to use ks a better question, and I don't have an answer (for LLM LoRAs, at least.)