Open-Llama: Complete training pipeline for building large language models
github.com
github.com
Feedback for /u/bayes-song - it'd be great to have a more info on the model card on HF - right now it's unclear the parameter count, # of total tokens you're planning on training on/how many you've trained on so far. An Evaluation section (maybe using lm-evaluation-harness) might be good as well?
But for now, you would need a good privacy reason to go this route.