Hey Jason, checked your HN bio and I don't see a contact. Found you on twitter but it seems I'm unable to DM you.
Went ahead and uploaded the image here: https://imgur.com/pJlzk6z
Went ahead and uploaded the image here: https://imgur.com/pJlzk6z
Least cost routing of prompt response. especially if time-to-respond is not as important as precision...
Also, is there a time-series ability in any LLM model (meaning "show me this [thing] based on this [input] but continually updated as I firehose the crap out of it"?
--
What if you could get execution estimates for a prompt?