The endgame for cameras is having no camera at all
theverge.com
theverge.com
As an image production apparatus, you can replace photography by machine learning (or really any other image creation process, drawing, painting, raytracing, whatever), and in these cases yes, photography becomes a little less important.
However, a photograph is not only an image, it is a trace of actual photons that existed and imprinted a photographic material. When you look at a photograph, you're not only looking at a picture, you're also looking at the imprint of reality. It's the same emotional effect as looking at your kid's hand print on paper, or wearing your grandmother's ring. There is affect involved that comes from the reality of the experience.
A souvenir picture is not only an image that helps you remember the time when you and your friends were doing something, it's an actual fossil of that moment. And that's what matters in photography. And that's why it's not going away any time soon.
What if you replace "photos" with "music", "art", or "literature"? Sure, we will get to a point where a computer can write a symphony that fits all the characteristics of Beethoven, or write a stylistically accurate Shakespearean sonnet. But it won't have the weight of the artist's observations behind it.
"I've noticed the many photographers here, [...]. Always the same conventional eyes, noses, mouths, waxy and smooth and cold. It still always remains dead. And the painted portraits have a life of their own that comes from deep in the soul of the painter and where the machine can't go." - Van Gogh
We've been asserting souls for quite a while now and we've been wrong every time. What does the weight of the artist's observations mean? How is that different from the result of machine learning, which is nothing if not carefully considered observations.
Note: The most expensive Van Gogh paintings sell for about 2 orders of magnitude more than the most expensive photographs.
(I was actually surprised at how much the most expensive photographs sold for)
I don't expect any reasonable answer to this, but...
There are two sets of photos, both exactly alike, yet one of those sets was created by capturing photons while the other was created by computers. Both were created by pointing a machine at your kid during a trip, and subsequently giving you a digital image. How to you tell those apart?
How do you know it's your kid on your kid's photos anyway?
1. Train ride with half hour viewing time to observe the Cristo Redentor. No cameras allowed.
2. Train ride to the statue for a photo of you in front of it, but while the camera can see both you and the statue, you can't see it from your position.
3. Green screen photo of you artificially in front of the statue.
I think a lot of people would, if they thought about it, strongly prefer the first option and ascribe little value to the others. But because they do have cameras and there's an expectation that you take a picture as a tourist, they end up viewing their trip through their camera lens or smartphone screen.
I think it would be a compelling sell if the camera app on a smartphone could prompt users with an offer of a professional-quality photo or video (optionally with the user green-screened in) of a concert, sporting event, or landmark. "Studio X has a professional recording of this event from N angles with better equipment than yours. Add this video to your library instead?"
The killer use for this might be to record/sync school concerts, plays, and sports. So many parents with mediocre cameras all trying to capture the event, and no good way right now to share the output.
I was also hoping for something more ambitious, such as "once you can detect the full electromagnetic spectrum hitting your phone from any angle, you won't need a lens to reconstruct a focused picture," as mentioned by blixt.
At a minimum, I was hoping for conceptualization and theorized advancement of the current display-pixel-as-camera-pixel train of thought.
But then it turned out to be a dishonest clickbait title with an article where the author wasn't really thinking clearly or with any sort of deep perspective at all about what cameras are for, or the value of what photographs of actual messy reality deliver.
Our memory is fragile and our eyes easily deceived. I'm hopeful this technology is introduced to the public first, so other institutions aren't able to abuse it before most people understand how untrustworthy pictures are.
Why do you think magazines use photoshop so heavily? People don't want the hot girl spare tires. Then want the perfect hair and big boobs.
It's all fake, but it's fake that is requested and voted for with money.
The problem here is not the tool, it's what kind of society we are building.
I don't disbelieve for a second that the expressiveness and realism of automatically-composed images won't have a huge explosion and become part of popular media. But I think there will be very long period during which some datasets will be much more salient than others. Think of how many orders of magnitude photos we have of dogs and cats compared to frogs and yaks... can they take the camera off of a product that can't image my friends petting a yak? I'm very curious when is enough, because the human ability to contrive absurd scenarios is frustratingly extensive, and unique "outlier" moments are exactly the ones many people want to capture.
Also I think I _have_ seen a similar art project like this, where the artist took a photo album of Paris(?) and gave people "cameras" that just identified the closest image location by GPS and would just keep regurgitating that image no matter how many you took. I can't seem to find it now...
You mean like movie computer graphics?
You're right that the idea isn't a good fit for instances like the ones you described, but I think it certainly has applications. For one thing, you could potentially synthesize an image of a part of a city from a particular era taken from a point of view that no actual photograph was taken from. Instead, it's a synthesis of contemporary photos with inferences made from contemporary weather data, contemporary maps, potentially drawings, etc. It's an exciting idea.
I see no exciting uses of this technology for recreational purposes. You wouldn't be able to take a photograph of the Library of Alexandria because that existed before we started gathering the data. And taking synthesized photos of your trips would make no sense, why would you even take the picture at all in that case?
My point is that it's a bad article. Technology is cool though.
Looking at my favorite photos I've taken, there are:
- Some photos of Iran (Google, being a US company, wouldn't operate there)
- The inside of a crashed plane
- A closeup of a dying hummingbird
- A laser that I use personally
- A boat adrift in the middle of the sea
All of these moments couldn't have been taken in this method, even with a boatload of content-aware scale and other people's photos stitched together.
For example: Banksy just tagged a building downtown. You rush there to get your picture in front of it. You can't take a photo of yourself with it because it's not yet in any of the databases used by your quasi-camera to stitch together photos.
It seems like the world is too variable for something like this to function practically for any but the most boring photos.
Five years later, you snap another photo of Banksy's latest, with an even lower-res camera. Two possible outcomes:
1. You've violated Banksy's copyright, whether he wanted that copyright or not, and the offending image is not present in your photo. You may or may not receive an automated warning.
2. You've recorded a crime, the authorities are notified immediately as your photo processes, and the image is again not present in your photo so as not to encourage the proliferation of graffiti.
The endgame leaves you with no camera, no one cares about Banksy who's in jail by now (with your help!), and why are you so concerned about violating copyright and encouraging crime, anyway? Why should technology enable you to do any such thing?
Maybe I could carefully explain to the software, "there's a male robin in this exact location and pose on the deck" and it could generate one, but why would I want to do that instead of just having a regular camera?
What does this do that I can't do with a camera? It sounds like a way more complicated, processing-intensive method to achieve an inferior result. It's a solution in search of a problem. It's not like decent cameras are prohibitively expensive, seeing as we've managed to include one in basically every phone now.
Most of the pictures I'd want to take are of something unique. Software may be able to figure where I'm standing, what permanent features are nearby, what the weather is like, what the people I'm with look like, but it's not going to know the transient features, like there's a lady with a weird looking hat in the background that I wanted to capture. So, what's the point?
You're right, what the author here is suggesting is not something you would want. He's imagining a magical omniscient fantasy technology, the closest approximation of which would leave you disappointed and robbed of agency as a photographer.
I propose, instead, you consider the elements he stitched together to reach his delusion and whether or not they could still lead to a very similar, regrettable future.
But that's not really "photography" (it's to photography what miniature plastic Eiffel towers are to architecture).
A much more interesting and much more futuristic approach would be a device capable of recording a whole scene without a conventional sensor and lens; something that would record photons not because it's hit by them, but because it knows where they are.
So this device would record all the light waves's position in a scene, at a given instant (the instant the image is taken), and then it would let you later reconstruct / produce any image from any position in the scene, with any kind of focus or bokeh, or whatever.
It would also let you walk into the scene like in a real "mannequin challenge", etc.
By modern definitions, the Holodeck clearly has the functionality of a 3D modeling tool and game engine: the crew frequently uses it to build VR prototypes of objects, situations and entire entertainment experiences.
It doesn't offer anything that we would recognize as modeling or animation tools, though. Instead the user describes objects and settings using any degree of precision -- "a steel table 3 meters long", "19th century London" -- to get a starting point, then iteratively adjusts the result by instructing the system.
(Nobody takes selfies on the Enterprise either. I guess they've lost their appeal when you can always just ask the computer: "Show me myself and Geordi smiling against a wall when we visited the Klingon High Council last year.")
So what's the endgame for selfie-takers?
For that matter, if we can simulate the lighting, etc. of an object, to the finest detail, why do we need the object itself?