Here is an idea, translate the text into a picture. An API will take text and generate a representation of the description using images from the web via google image search. Here is an idea, translate the text into a picture. An API will take text and generate a representation of the description using images from the web via google image search.WordsEye constructs 3d images from text descriptions http://www.wordseye.com/
Sketch2Photo takes a crude annotated sketch, and creates a composite image: http://cg.cs.tsinghua.edu.cn/montage/main.htm
The output of wordseye isn't great - it looks a bit 90's POV-Ray. Sketch2Photo does a nicer job - more like automated photoshopping - but needs more assistance on the placement of objects in the scene.
I'm sure there's others.