Walk through a 3D model of Y Combinator
matterport.com
matterport.com
This'll work great for navigating buildings—I can imagine it being a hit with real-estate sites—but the hiding of detail in the 3D view suggests the tech wouldn't work as well for other applications, like VR. (I'd love to be wrong here, though!)
As for VR... we're playing around with it a whole lot, and it actually looks really great! Hopefully it'll get in front of more people somehow soon.
This is a cool thesis. I would definitely like to see how this could work in VR.
Duplicate of https://news.ycombinator.com/item?id=8524360
TBH I just want to know when it auto-compiles into Doom WAD files (or the modern equivalent) so I can run through famous buildings shooting monsters.
Saying "Our model quality has changed a lot in 2 years as well." and then presenting the 2014 photo is pretty sketchy. An accurate comparison would be viewing the model from the same position as the photo, using the model, not the photo. Or allowing arbitrary angles - I bet if you look under the machines on the front desk on the actual model, they've melted into the desk a little bit.
Or - I do understand the quality is not there yet - say 'Our viewing experience has changed a lot in 2 years as well.' and show the old viewer and new 2D/3D combined viewer.
I think 60 frames per second would be smooth enough. Walking speed 5km/hour is roughly 1 meter/second. Per frame distance then is about 0.02 meters.
Looking at the map the YC office seems to be about 30 x 40 meters. Imagining a grid with lines each 0.02 meters overlaid on it, it seems you would need 3000000 panoramas. That might sound way too many, but wait! We live in the future.
One spherical panorama of 10000x5000 pixels seems to be about 4MB jpeg-compressed. So you would need only 12 terabytes of space. Also since you need 60 of these each second, you need storage that can move 240MB/sec, which is lower than the current speeds of SSDs.
1TB SSD seems to cost about $400, so for only $4800 you would have enough speedy space to store the panoramas. Enough space to explore whole YC building with no snapping at all between frames, with complete realism. Actually you could even do stereo 3D, as you already have the data.
Now it's a different question entirely on how we could take those 3000000 panoramas. Even if you had a Double Robotics bot with a spherical camera attached to it going around the space, snapping 10 panoramas per second, it would take 4 days to complete. While that itself is tolerable, how it would know its position and control its movement to the required accuracy I have no idea.
Still it blows my mind that storing all that would be possible.
also, it assumes you can only continuously walk through the space right? what about standing still and swaying your head sub-2cm?
Interesting thought exercise, thanks.
Paper: http://research.microsoft.com/en-us/um/redmond/groups/ivm/iv...
Web Page, Video, Slides: http://research.microsoft.com/en-us/um/redmond/groups/ivm/iv...
Quoting: "Whereas many previous systems have used still pho- tography and 3D scene modeling, we avoid explicit 3D reconstruction because it tends to be brittle."
I appreciate it could just be a personal thing, I've never asked anyone else what they think, but I seem to regularly accidentally look up because click to location often competes with viewport manipulation, you then can't get it 'level' which is mildly OCD annoying and the controls also feel backwards as you have to pull right to go left, pull down to look up, etc. Also the sensitivity of up/down seems exaggerated to left/right, but probably because it's windowed, I actually don't know.
I can't describe very well what's wrong, it just generally 'feels' wrong.
Google aren't exactly known for their UIs and as far as I can remember that control system was their first go at doing it and they've never changed it.
Ultimately he'd want a drone to fly around a house and automatically creaete a 3d model and 2d plan. That would probably be exceptionally useful. He's pretty efficient at doing measurements and drawings, so fully automatic is almost the only way to be useful to him. Of course that's just one guy but I thought the example may be helpful.
The model is freely-available (CC0) in FBX ("full geometries and color textures including interiors") and STL formats. http://www.cooperhewitt.org/about/mansionmodel/.
http://www.3dsystems.com/ did the scanning/photography.
In video games like GTA, LA Noire and Watch Dogs real cities are mapped out using a sort of "conceptual compression" that leaves landmarks but somehow brings the space closer together. My sense is that this is a labor intensive process, but what if it could be automated?
It would be interesting way to explore a place, though with the obvious pitfall that what's not included in the map "doesn't exist".
GTA, specifically, doesn't use all the landmarks they find - they instead use rough facsimiles that give the same impression. It's really rather brilliant - a different design, but somehow it's familiar if you've been there or seen that.
Ultimately, for video games, they are still designed by hand, because of the "if it's not there, it doesn't exist" problem and for a host of other reasons. Unfortunately, for space reasons, most of it is non-interactive in a meaningful way.
There's some interesting work being done by ESRI that could hopefully lead to virtual city designs that are almost fully interactive. Imagine GTA where every building could be entered and every object could be interacted with because they are generated on the fly.
Sort of a realistic Minecraft. Very sort of, but still.
See this popular work from 2009, "Building Rome in a Day": http://grail.cs.washington.edu/rome/
Quoting: "The data set consists of 150,000 images from Flickr.com associated with the tags "Rome" or "Roma". Matching and reconstruction took a total of 21 hours on a cluster with 496 compute cores. Upon matching, the images organized themselves into a number of groups corresponding to the major landmarks in the city of Rome. Amongst these clusters can be found the Colosseum, St. Peter's Basilica, Trevi Fountain and the Pantheon."
I thought we were in the 21st century :/
https://mega.co.nz/#!e9EmiaqK!89u5YTkFFGHFxpCVFNZnE22uetkLBY...
Is this the Highest resolution the model and textures possible on the camera currently?
Is this photogrammetry?
If it makes an obj, why can't I move arbitrarily?
How does it compare to AgiSoft Photoscan or 123D catch?
Don't know Photoscan or 123D catch enough to comment spesifically, but in general many other 3d companies focus on scanning small objects or features, while we do large buildings/indoor spaces/rooms better (faster/better quality/cheaper/more convenient) than anyone else I've seen :)
That said, Photoscan's UI is incredibly poor, and the software has bugs (particularly around CUDA) so I'd be interested in alternatives.
123D catch (from Autodesk) does cloud based processing like you guys.
I've lost touch with architectural 3D "capture" - who else is working on this space? Are there any consumer-level offerings?
On iPad it said "upgrade to iOS8", on W7/FF33 it said "Oops, something went wrong", on W8.1/FF33 it went all the way through the loading and then got stuck with just one pixel in the progress bar left to go. :-|
http://i.imgur.com/lNKLbj5.png
Typekit errors are due to the Referrers being blocked, this is very common and never disastrous. Disqus and Google Analytics are just blocked at the domain level.
"3 free models per month"
so the camera by itself (+ whatever software) doesn't just spit out a point cloud i can do whatever with?
http://i.imgur.com/ctPgj2v.png
Chrome 38.0.2125.104 on a mac.