AI movie posters
noahveltman.com
noahveltman.com
For the ones I did know, I almost always could see a similarity to the "feel" of the movie. But the Rear Window one?? What happened there? Didn't capture the "feel" whatsoever IMO. Of course there are windows but it looks more like a backalley in Japan. The color palette is way off.
Tangentially related, this AI image generation account on Twitter never ceases to fascinate me. Many of the photos are seeded with an initial image however, which reduces the magic/specialness for me a bit: https://twitter.com/images_ai
Nonetheless, this was super fun.
For example the very strong color scheme in Miami Vice, which is basically 100% something that emerged in popular culture after the movie.
So when the AI is generating the pictures based on the plot, could it be deep down generating the pictures partly based on how we people remember the movie?
Not that this should matter in this context, and no idea how common or uncommon this particular "skill" is, but I can often identify a movie that I've seen years, even decades ago, by seeing just a few seconds of it. I guess this is a form of photographic memory. But, it's not only the raw snapshots, instead it's a mix of image quality, content and if I catch a glimpse of it, the story and characters too. Is there such a thing is "moving pictures memory"?
Edit: I did identify "Being John Malkovich (1999)"! Then again, his face appears in the picture.
No idea either. But add one more if you’re counting, it happens to me as you describe.
Or used to. Television was an important trigger because it shows movies unprompted; switching to alternatives where I have to consciously pick what to watch removed the opportunity for the effect to manifest.
The last one, Star Wars, I think needs some deep inspection to get it and is missing the Vader ( which I think anyone with identify with the movie)
The ones that I could identify: * Star Wars * Edward Scissorhands * John Malkovich * National Treasure * Hot Tub Time Machine * Willy Wonka * Space Jam * The Hurt Locker * Midsommar (this one gives me flashbacks, btw) * Mad Max * Fear and Loathing * Oceans 11
Everything else was too obscure. Curious which ones other people could identify.
Willy Wonka & the Chocolate Factory in particular seems to be a direct reconstruction of features from that movie.
I immediately got a feeling of what the movie is about, but a lot of times I named the wrong movie or couldn't remember it. The feeling was similar to when you have something at the tip of the tongue..
Some were obvious after reading the title: Star Wars, Harold & Kumar Go to White Castle (never saw it), Being John Malkovich, Space Jam (never saw it), Miami Vice (never saw it), Office Space.
Some that I half guessed or felt like I knew the name of the film but could not immediately recall the name: Castaway, Mad Max: Fury Road, Edward Scissorhands (Thought it was The Nightmare Before Christmas but Tim Burton so close enough)
Obvious ones: Willy Wonka & the Chocolate Factory
And I don't think its simply because of the likelihood of guessing a popular movie, as many of the titles I've never heard of. Although some are a giveaway by their content and focusing on an obvious subject - Ghost Busters has clear ghosts; Hurt Locker has a bunch of bomb-disposal guys. Others like Edward Sissorhands, have no giveaways but just embody the style; I'd love to see the descriptions provided to the AI
Unless the descriptions were particularly detailed, I would expect that a lot of this comes from the training data and the descriptions are just prompts for the model to recall which film it is.
For instance, Willy Wonka & the Chocolate Factory (1971) is clearly based on this real poster:
https://www.amazon.com/-/es/Póster-Online-Willy-Chocolate-Fa...
This seems less “generate a poster from a description” and more “recall half-memorised posters“.
VQGAN does the generation work, CLIP just says if it's good not, improve the latents, repeat. Here's a good technical writeup: https://ljvmiranda921.github.io/notebook/2021/08/08/clip-vqg...
And of course, the most-common VQGAN was trained on ImageNet, which likely doesn't have every movie poster as training data. (it could be in CLIP though)
> a brief text description of a movie
However apart from the existence of a golden ticket, I wouldn’t expect those details to make it into a brief description of the film. And yet there’s an original poster matching those details that the VQGAN + CLIP generated image seems to draw from.
That's not enough for reconstructing the face of John Malkovich from text, you need minute facial feature parameters (eye shape, nose shape, eye-nose distances etc etc)
Star Wars is an interesting example because it appears to include elements lofted directly from the film (bits of stormtroopers body) alongside a princess who definitely isn't Leia. The algorithm might be creating things from scratch at a high level, but the constituent elements are pretty clearly close reproductions of parts of the source material
The interesting thing to do, I think, would be to have data set of general images, and use a movie description to pull from those images.
When I looked at that one, I actually "saw" a figure holding distorted scissors imagery.
Trippy stuff!
The matrix I answered as John Wick, which is pretty interesting.
Fear and Loathing was the other I got but I only tried 10.
The way people's brains are analyzing and picking apart these images differently is fascinating to me
I'm pretty impressed and amused by this.
Thought process: "Making out on a beach? Oh that looks like an Italian flag!"
I also missed the personal connection and the prompt I get from seeing human faces, actors I like and enjoy.
Yes, I think I detect a definate bias towards a Picasso/cartoonesque style :-)
Even Dr Stangelove is in colour.
I think our jobs are safe for the forseeable future.
https://thorsbyprojects.tumblr.com/post/660029925870518272/a...
It's clearly impossible to end up with a white haired man and a Delorian from the words "Back to the Future", so it likely has way, way more information about the movie than the title. Possibly even movie posters. I think it's a bit misleading.
I'm pretty impressed by ones like Ferris Bueller's Day Off having what I'm pretty sure is the Cubs logo and Wrigley Field, plus Bill and Ted's Excellent Adventure having multiple phonebooths.
What does that mean for the philosophy of Art? Art is meant to communicate emotion [0] but how can something that has no emotion communicate it? And if we exclude this from being Art because of that, then how do we tell "real" Art from this, when there's nothing to distinguish between them?
What happens to Art when you can get an off-the-shelf ML algorithm to produce original, creative, works like this on demand? And what happens to commercial artists and illustrators once we train the algorithm to do their job?
Though that's not all bad - Art could become a hobby done for the joy of creating, rather than a profession done to make money.
Interesting times.
Perhaps the same as what happened to artisans such as some makers of furniture, shoes, and bespoke cloting. They charge a premium for the "authenticity" of their products. Artisanal, single-source, small batch art.
Whilst it's aesthetically pleasing in a semi abstract kind of way and you can definitely imagine people paying to put it on their wall, nobody's going to confuse it for the actual human created film posters, so I think the people that make film posters for a living will be just fine (unless all future film directors and promoters want their film posters to be a collage of objects from the film on a psychedelic abstract background, I guess)
Decorative art creators might be threatened. But they'e been threatened by reprints for a long time. Political/philosophical art not so much as they are more cognitive and require the artist's societal involvement in some way (even as outcasts).
These AIs generate a specific style of art and you can put it up on your wall as you can put up a reprint of human art. But when they're used with purpose, they can become tools of human artists.
I'm no expert, but I couldn't tell this generated art from an actual artist (e.g.) "experimenting with abstract representations of movie themes". Could an expert tell them apart? Are there any obvious "tells" for AI-generated art, and can that be overcome with better training?
I feel like we're watching the equivalent of the first computer chess games, with people saying "computers will never beat a grand master because they don't understand the strategic complexity".
People still play chess, despite the fact that their phone can beat them every time. Maybe it doesn't matter that a computer is a better artist, either.
The system is garbage in garbage out. So, feeding it random noise inputs will give noise results. The human factor is “prompt engineering” the inputs to try to coax the elephant in the desired direction. And, in evaluating the outputs.
Sometimes it feels like asking a child to draw something for you. I spent some time trying to get “storm clouds made of lava” but the child insisted “No, dad! Lava stays on the ground. Like THIS!” https://www.reddit.com/r/bigsleep/comments/o033x1/playing_wi...
Similarly, while trying to make a “tempest in a teapot”, instead of a storm, the child was a fan of the My Little Pony character Tempest Shadow. https://www.reddit.com/r/bigsleep/comments/o04sin/trying_to_...
Do you think this will change and get better? Would it help if the AI was capable of interactively refining the inputs (I guess by asking questions instead of making assumptions)?
In theory, given enough machine power, I could see a branching, interactive interface. Like if there are multiple major hills in the latent space ("Tempest-> a storm cloud" vs. "Tempest -> a cartoon pony") the system could identify them all and generate a set of images with different biases towards each hill. Then the user can pick which direction to climb.
The current system on high-end hardware takes about a minute before results for one image even begin to be recognizable. And, about 3 minutes before you can be somewhat confident how it's generally going to turn out.
Just like chess AIs are the result of human work by the way. I don't know if there is a league of AI chess, but if there were, you'd always have the developer teams behind them who make strategic decisions on how to develop the AI. They become chess players of another kind but chess players nonetheless.
I have never picked up the book or watched the movie Fear and Loathing in Las Vegas. I probably couldn't have told you anything about them beyond what was included in the text description the AI got. Yet I instantly knew that was the movie.
I could even envision a few of these as art work for a special edition blu-ray or something.
It's very fascinating to watch and almost empathize with the computer as to how it joined these concepts together (i.e. "bullets fly" ends up like a kind of bullet-hummingbird type creature)
There is something sort of magic about these images. The satisfaction I felt when I saw a picture and instantly had a thought that turned out to be correct was really refreshing.
It's not the same feeling you get when you put a lot of effort into solving a problem and eventually getting there, but more of an artistic sense. Interesting.
These images make me wonder what the "brief text description" is for each.
SPOILERS BELOW
I got most of them by looking at either landscapes or color palettes. In some cases, certain characters were actually visible in the collage.
Wizard of Oz, Ghost Busters, Fear and Loathing in Las Vegas, Mad Max: Fury Road, Office Space, Space Jam, Willy Wonka and the Chocolate Factory, National Treasure, Star Wars.
var answers = document.querySelectorAll('.button'); answers.forEach( answer => { answer.click(); });
I would have appreciated some more recent ones though. I don’t quite remember all movies from the 80s.