Nvidia Canvas: AI for creating realistic landscape images
nvidia.com
nvidia.com
They could turn Canvas into a web app and charge a monthly subscription. Alternatively they could go the OpenAI GPT-3/DALL-E-2 route and give it away as a web app or API to generate a huge potential customer list. Instead, they're only interested in technology demonstrations.
I'm not arguing that this is a good _or_ bad thing. It's just interesting to watch a company drive one of the greatest innovations humankind has ever developed (AI), yet fail to capitalize on the resulting value creation due to a hardware focused culture.
I'm quite lost on what you think would pay for this kind of software?
It's the reason why Apple no longer charges for OS updates. Nvidia has essentially the same business model.
It's very hard to sell software to $averageconsumer unless that software is "free"
I don't mean to denigrate this, the results are clearly interesting, but I just don't understand what problem this solves, it just seems to raise the noise floor on reality.
At some point, there is going to be some sort of higher level research for ML in terms of generating an architecture for a particular task. And all this research is going to be used for this.
The training labels may have been “segmentation maps”. These are regions of an image with a known scene description such as “cloud”, “trees”, “sky”. I’m not certain what model they use, but I bet it is a Stylegan2/3 modified to generate an image from a given set of segmentation masks.
Indeed, without the research context, it’s a little strange “why” you would want a product like this. Nvidia has done a lot of research to get GAN to run very fast on their RTX cards due to being mostly convolutional, operating directly in pixel (or wavelet) space rather than an embedding space. On my RTX 2070, I can run Stylegan2 at 1024px at a somewhat reasonable 10 FPS.
I'm the founder of https://ayvri.com, and we have a 3D virtual world where outdoor athletes watch their activities, and the activities of others.
As the resolution (and speed) of our 3D world improved, people got more interested and engaged with it.
I believe this is the future of video. Not volumetrically created through 20+ cameras, but with a single camera capturing the scene, and AI filling in the blanks based on what it knows.
Even the most computer illiterate of people these days are able to scribble an MS Paint landscape and have it (usually) turn into a gorgeous seascape or mountain vista.
First program/app since maybe WordLens (that old iOS real time translation overlay app) that gets consistent “wows” out of virtually everybody.