This looks super cool! Is there currently a limit to how big a repo can be for this to work efficiently?
We stress-tested with repos like langchain, llamaindex, kubernetes and there the retrieval still needs work to effectively return relevant chunks. This is still an open research question.