Great start!
Are you planning to add to following:
Retries w exponential backoff, Caching, Streaming output, Function-calling support
Retries w exponential backoff, Caching, Streaming output, Function-calling support
Streaming output and function-calling support is interesting
why would you cache the openai call instead of the endpoint that's receiving the user call?