Long-form factuality in large language models
arxiv.org
arxiv.org
[1] https://editor.factiverse.ai [2] https://arxiv.org/abs/2402.12147
Still needs work
[1]: https://arxiv.org/abs/2305.14251
[2]: https://github.com/shmsw25/FActScore?tab=readme-ov-file#to-u...
I am not an expert in this research but this seems like this is just a slippery slope all in the pursuit of cost, first and foremost.
Given that the summary focuses on cost and this paragraph mentions cost as the first point, it sure seems like these folks only goal goal is to just take humans out of the mix entirely when it comes to facts.
Is this a good idea? I am not sure.
You can't trust that this will work on your knowledge domain or that it'll work in the future.