Excited to give this a try! A visual environment seems like a natural fit for chained LLM calls.
Have you considered supporting OSS LLMs? Inference servers like LocalAI or vLLM expose APIs for various OSS models with OpenAI-compatible endpoints, so might not be much more work to integrate.