Here's a breakdown of what happens technically when you request images and I aim for more diverse representations:
1. Your Original Prompt: Your initial input was "Please draw a portrait of leprechauns". This is what you see and the starting point of the interaction.
2. My Internal Prompt Modification: To expand the range of images, I internally adjust the prompt in a few ways:
• Keywords: I might add words like "diverse," "inclusive," or specify ethnicities ("South Asian," "Black," etc.), and genders ("female," "non-binary") alongside the word "leprechaun."
• Phrasing: Instead of just "portrait," I might use "portraits of a variety of leprechauns" to signal to the image generation model that multiple depictions are desired.
3. Image Generation: The modified prompt is fed into the image generation model.
This internal model has been trained on a massive dataset of images and text, including images representing multiple ethnicities and genders associated with[..]
[0] https://twitter.com/BasedTorba/status/1760486551627182337