Visual Turing Test
turing.deepart.io
turing.deepart.io
For me, the giveaway for computer-generated images was that they don't focus on the "subject" and have too much detail on the background.
For example, look at the computer-generated boats - they blend into the background. It's hard to tell where the boat ends and the rocks begin, or where the sails end and the clouds begin. Or look at the grey house in front of the dude's face or the pipe behind the glasses lady's head, which bring in too much attention but add nothing to the painting.
I had read headlines about computer-generated images but had never seen them, yet I managed 9/10, anticipating that an artist's technique is not just relation between color etc. but a philosophy.
Edit: Also, it is not a turing test in true sense, since it's one sided. It would be turing test if I tell them what to draw & then it's tough to distinguish between what an artist & computer paints.
[1] http://turing.deepart.io/t/4.jpg
"one is painted by a human and another one is generated by artificial intelligence based on a photo and a style of a painter."
I find this to be easily "beatable" by simply judging which one is most likely to have its origin in a photo (vanity, and such) with a fallback on the one with a lot of repeating patterns.
I don't think programs written to imitate a craft, or even to learn to imitate a craft from examples, count as AI, no matter how impressive.
I disagree - nobody would say drawings/paintings (by humans) based on photos or observations by eye are less legitimate than scenes completely imagined.
Interestingly, this is a Turing test that, when applied purely to humans, doesn't require said humans to share a common language.
I have to quibble though that the Turing test is meant to be a test of intelligence, and this sort of task seems pretty different, although it may actually be visual-AI-hard (to invent a new term). So may deserve a different name.
Those photo modifications kind of fit into that theme: if you can't distinguish the filters from real art, they passed an indistinguishability test.
But the opposite heuristic doesn't work either, because not all the computer paintings are abstract. It mimics a lot of different styles and I am extremely impressed.
I too used the same method and 5/10. Looking at the other comments, it seems people have better idea what art is than I do.
I am convinced intelligence programming will tackle these issues some day. But apparently not yet.
also the high-fidelity, near-photographic paintings are always human.
I suspect that we can be trained to spot fakes too to make it harder. Perhaps there will be some co-evolution?
Somehow I cannot believe how fast they got the idea to create profit from it: http://deepart.io/
As a side note, for me the pace was one of the reasons to move from science to data science (shameless plug: http://p.migdal.pl/2015/12/14/sci-to-data-sci.html).
If I wouldn't know that some of them were generated by ANN, I wouldn't question that like 90-95% of them were painted by a human. And I consider myself pretty alright at painting.
These 2 are especially awesome:
1. The photos are low-res enough to hide the most ridiculous artifacts produced by neural nets
2. The machine-generated images are trying to emulate the likes of van Gogh, not the likes of Leonardo (again hiding the extent of NNs complete inability to understand what they're doing)
3. Most people simply neither paint nor appreciate visual art, so this is not as powerful as the Turing test! (One point of the Turing test is that a human easily passes it, because all humans speak; they don't all paint.)
Prepare for some people actually getting angry about it once you tell that the paintings they liked most are computer generated (from pictures but still).
My girlfriend is a fine arts major and isn't speaking to me right now.
The computer does not duplicate selective detail. Artists will put detail in areas of focus, and intentionally obscure or abstract other areas. Also, there is often an interplay between medium and message. To that last point, I'm guessing there is still a human "artist" who chose the reference scene and pose and artist style, even if they didn't paint it directly. Towards that, I see this as a creative tool to be included in some future Photoshop, or similar.
Also... Share on Facebook? "App Not Setup: This app is still in development mode, and you don't have access to it. Switch to a registered test user or ask an app admin for permissions."
These guys seem to be trying harder to actually make this a Turing test, but as the commenters above have noted, the problem with using source photographs makes this a fairly imperfect test. I have a suggested fix for that.
In my last effort in that post, I actually used one source photograph, and then compared the output by a human painter and the AI painter. My original results were terrible, but then quickly vastly improved by a better implementation of the algorithm, the Deep Forger implementation.
I think that this variant in the procedure, having both the human and the AI start from the same source photograph, could make for some interesting variants on Turing testing.
I wonder if the algorithm could be improved if it had a sense of what is "important" in the picture, and then choose a different algorithm to process important and unimportant portions.
Right now neural nets do pretty poorly with "loose/messy" styles (you get noise instead of brushstrokes; if they posted high-resolution images there, I cannot believe anyone would score less than 10/10.) Neural nets do exceptionally poorly with "neat" styles (try copying the style of the Mona Lisa into your photo and it'll paint an eye into your mouth.)
However, I'm pretty sure that an algorithm hand-crafted to replicate an artist's style might fool me (just like a hand-crafted chess algorithm wipes the floor with a neural net plus some tree search; not so for Go, I know - I don't even play Go so I don't have a firm opinion on that one.)
I must say that this whole "photo + style" thing really gets my goat - not the research as much as the sort of press coverage it gets. I project that a stock market crash will improve such coverage significantly.
One thing I did find interesting is that I did the test initially and got 4/10 (most felt like guesses to be honest). I repeated the test before looking at the answers (though obviously knowing the score introduces some bias), but zoomed the page so the images were about 2-3x bigger than the default. My score increased to 9/10, and I was much more confident of my answers whilst selecting them.
- Is the style of the painting applied uniformly over the whole image? Humans will, for effect, leave parts of a painting less emphasized and developed, whereas this algorithm will generally apply a style to the entirety of an image.
- If a painting has a surreal style, the subject is generally also surreal. Human painters distort shapes and forms. This is not done by this algorithm.
- Humans will add contrast to make objects stand out, even if colours are similar to their background. This is something the computers haven't yet completely figured out.
Still, this is a very impressive algorithm.
Are these computer-generated images based on photographs of real subjects, or on photographs of stylized paintings?
I don't think this test is very interesting until we can see the images the computer-generated ones were based on.
It's amazing to live in a time when questions about human identity and expression are more than just hypothetical.
- obvious artifacts in the neural networks
- humans paint light differently than how a camera captures them, evident even after heavy processing.
I do try to appreciate art, but I'll readily admit that I just don't get a lot of it.
I guess as of today T-1000 would get his bitch ass kicked.
First, all images could've been painted by a human.
Second, if I always click the left one, it says Your result is 5/10. A turing test is to test "a machine's ability to exhibit intelligent behavior equivalent to, or indistinguishable from, that of a human". If I – presumably a human being – score 5 out of 10, the test is not a Turing test after all, because humans aren't tested in a Turing test.
Edit: My observation was obviously wrong. Thanks for pointing it out.