github.com
https://arxiv.org/abs/2403.01081
Would be curious to use same dataset without fine-tuned llm for RAG over same data.
Then one could immediately make use of building the dataset and then measure gains from training.
Reply on news.ycombinator.com