Show HN: Visualise novels using Midjourney and GPT-4
parsabg.com
parsabg.com
I had an idea to use RAG to extract all relevant descriptions of given object and compile them to a detailed description. The description would be fed to a text-to-image model.
Have you considered something similar? It would be harder to implement, but the results could be more precise and it would be possible to cover books GPT-4 is not familiar with.