davinci-003 has a 4k tokens cap, you cannot just 'shove all your information in the prompt.' Can parallelize but then rate limits become a problem. Really curious to hear how these guys are handling it.
https://langchain.readthedocs.io/en/latest/modules/chains/co...