Just got doxxed to within 15 miles by a vision model, from only a single photo
twitter.com
twitter.com
You might think making an AI model will be cheaper, but if you commoditize it like this the additional police work from stalkers will outweigh any savings by multiple orders of magnitude
1) Easy mode: Cityscape of Naples, Italy; Bot 100% match, nice, I guess.
2) Mid-mode: A road in Glenoak Hills, CA; Bot says Malibu, CA... Never been there.
3) Hard-mode: Main lake in Ostróda, Poland; Bot says it's in Manhattan and that the vegetation matches Northern American :-D Ahahaha, didn't even get the continent right so...
Note: Tested on personal photos, with EXIF stripped, not widely/at all uploaded to the nets.
So... perhaps works great if you feed it back stock photos you've grabbed from the first result of a google search? Because my mileage is barely an inch.
https://www.youtube.com/@rainbolttwo/featured
AI repeatedly winning against him in a head-to-head:
Edit:
From the rainbolt interview, they claim trained on 200k images, 92% country accuracy with median error of 44 km... median error is... not remotely what I'm experiencing, but the AI seems to be doing great job in geoguesser match vs rainbolt. Also supposed trained to be aware of cardinal directions.
That it guessed this one correctly is a miracle to me.
Then I tried it with a photo of an abandoned structure in rural France and it guessed UK because some words in a visible Graffiti were in English.
So, it is hit and miss.
I gave it a much harder photo from the Borrego Badlands, and it got SoCal, but was 200mi off.
Some of them (pictures of a small town beach) were described as:
"The photo was taken from a boat in the XYZ Harbour, looking towards the city of <Large City Name>. The distinctive shape of the <literally tallest building in the country> Tower can be seen in the background. The ferry terminal is also visible on the left side of the photo."
needless to say, the tiny town in question is about 300km from the city mentioned, it does not have a single tower (I think it may be looking at a traffic pole?), and does not have a ferry terminal (there was a boat in the picture).
While it correctly guessed the country and general elements in the scene, it failed to pick up on those towers and the nuances of the terrain, which led it to suggest a very different place hundreds of km away.
Still, doesn't make me less wary of posting images from my neighborhood online... it'll only get better.
I guess that this is really hit and miss depending on a bunch of parameters.
I wonder how Street View would perform on signatures such as Shazam uses
Chicagoland park-> guessed central park NYC
Rural Nebraska -> guessed Iowa
Shoshone Wyoming -> guessed Glacier Montana
AI: Scottish residential houses, and a hill. This is clearly Glasgow.
https://www.reddit.com/r/wherewasthistaken
Random streetview shots seems to guess countries correctly so far, or at least region with similar construction. But that still means being (confidently) off by 1000s of km.
I'm not sure. Isn't the consensus that neural networks get better, if you make them bigger and give them harder problems?
I mean… maybe… but that also makes them more expensive to run…
[0] https://www.nbcnews.com/news/world/stalking-suspect-allegedl...
People who believe "you can't dox from a picture of trees, because even tho conceptually trivial, you don't know how to search the database efficiently enough" are people who also do security thru obscurity.
When it comes to the fact that this is possible at all following various geolocation and OSINT accounts has had a quite sobering effect on me quite some times me ago.
I would like to point out josemonkey particularly, because he really explains the process very well. By and large I think it is first and foremost a fairly tedious task, but doesn't require any access to special services or special skills. Of course the vision model could change the tedious part.
Overall I think what the OSINT people do and what the vision model from the article simplifies is entertaining, but ultimately not super relevant, because reality is more like xkcd 538[1].
Most people have a device on them that in some way reveals their location more or less accurately. Location data in the grand scheme of things is a commodity for quite some time already.
While I think it is worth fighting that, I guess pragmatically we should all live under the assumption that our location is public information.
Sounds terrifying