3 karma · joined January 15, 2024
Now building with LLMs and multi-modal models.
This method uses SDXL and is supposed to be more consistent than previous known methods according to the paper.
We built something similar to query DB. Created two versions, one of which was agent based that had a maker-checker style of generation. Basically one generates and the other checks its correctness and if objective has been acheived.
The accuracy improves in the agent driven framework, at the cost of latency.
The main challenge for production settings
1) soon becomes optimizing the latency for each step and 2) dealing with complex queries at a reasonable cost (GPT-4 is bit expensive at scale)
We evaluated multiple options and settled on GPT-3Turbo in the short term. However we are not happy with the latency, especially if you create an agentic solution.