would love to see the data quality & diversity metrics. It's kind of hard for me to go through tens of thousands of examples to understand the quality of my data sometimes.
Just a slight bump, I've now added more support for unauthenticated users, so that people can check it out without having to sign in. This hopefully should reduce the barrier for people try this app out!
One thing I'm trying to explore is how AI can help accelerate scientific progress, so I spent the weekend hacking together a research assistant that summarizes the most relevant research papers from arxiv for a given query.
Right now it is only papers from the "cs.CL" category, but I do plan to expand on this a bit more (at least cover the whole cs corpus).