The whole point of this kind of thing should be to reward people who can recognize "that architectural style wasn't invented until the 13th century" but that's precisely the sort of thing image models cannot do reliably.
The whole point of this kind of thing should be to reward people who can recognize "that architectural style wasn't invented until the 13th century" but that's precisely the sort of thing image models cannot do reliably.
For example, consider this imagery from today's challenge: https://firebasestorage.googleapis.com/v0/b/fastab-f08e9.app...
These are some incredible monoliths: if they were real, I feel like I would have heard about them? And if they did... that's so cool. But because it's AI generated, I have a very low confidence level that this ever existed at all. Which is sad.
Which is funny, because the monoliths in the AI video look more eroded than the real ones today.
This looked like a nice idea at first glance. At second glance, it's really bad because you have to assume that everything you see in these videos can be wrong or misleading.
This is entirely possible, as the incredible accuracy[1] of non-generative picture location models (a very similar problem) shows.
[1] https://paperswithcode.com/sota/image-based-localization-on-...