Dude! This is so freaking cool.
Curious to know how much are you spending in LLMs to generate each model.
The complexity of the model does heavily influence the cost. Some models can cost me at little as 7 cents to generate. Others can be a couple bucks.
At the moment I'm focusing on user experience. I just trying to make a product people love as number 1 priority. I'm constantly trying to improve the speed and quality of the model outputs for user. Thankfully, optimising speed helps bring costs down as a side effect too!
I found building an MCP was the easiest way to avoid paying for inference or having to charge to try it out. Anyone can add the MCP to chatgpt or claude and use their subscription to do the inference.
[0] sawdust.diy