Seems like the search is based only on the transcript/dialogue - not an image embedding. Would be super cool to actually use some CLIP/embedding search on these for a more effective fuzzy lookup.
When I use Frinkiac, I already know the quote (or at least a few words from it), and I simply want a still or short gif of it (with a customized caption typically).
YMMV of course.
I’d guess famous characters like Bart and Marge and other Simpsons characters would likely be known by the tokenizer so it’d be pretty easy. So then you’d be able to guess.
Feel free to correct me on small details if anyone has this more fresh in their mind but I’m roughly correct here.