82 karma · joined January 2, 2015
This is Harsha, from the team at 3difyMe. We rolled out the first version of 3difyMe a week ago and wanted to get some feedback from HN community.
<TLDR>
3difyMe converts flat photos of T-shirts and other apparel into 3D models. You can generate multiple combinations by changing the {human, pose, background, footwear} and click unlimited photos or videos. These are ready to go to your online store or marketing platforms.
</TLDR>
—— Longer version aka Why are we doing this?
Circa Nov 2021, we were toying around with realistic 3D human models and wondered what business problems we could solve. Avatars were one (thanks to the incoming metaverse). But we didn't know if the timing was right. So we explored further and stumbled upon the E-commerce photography space. At the moment, apparel retailers typically use photos that are: 1. Flat-lay 2. Shot on a mannequin 3. Shot by professional agencies
Cost, effort, time and quality increase as you move down this list. A typical agency will charge you $15/product with a minimum order of 20 products and deliver in 2 weeks.
At 3difyMe we are trying to go from 1 to 3 using tech. No minimum order. And 1 minute to get results.
Here are our hypotheses about this: Set 1: > Smaller retailers doing flat-lay will be able to improve their visuals significantly with little effort. > Bigger retailers should be able to reduce spend, time and effort for a slight trade-off on quality.
Set 2: The ‘per image’ limit is irrelevant in a virtual studio. You’re free to take as many photos as you want. If retailers had a lot of photos, they could: > A/B test their visuals > Improve their marketing game with additional content at no extra cost
Set 3: In 2022, consumer attention is fully engaged in video (TikTok, YouTube, Netflix) and retailers barely have access to this medium. That’s because apparel video shoots are difficult and expensive. But not in 3D! > The next innovation cycle in retail is video and 3difyMe could make it accessible.
—— A note about the open source we’ve used: 3difyMe is built with Blender, ThreeJS and React. It’s just about amazing and unbelievable that these masterpieces are free and maintained so actively. We plan to contribute back both in code and funding in the near future. We ♥ you Blender and ThreeJS.
—
Feedback of all kinds is welcome. On product, market and product-market fit. :)
If you want to personally reach out, please write to us at harsha@3dify.me and team@3dify.me.
Does the sender have to be up for the receiever to receive?
In the best case, carbon removal and sequestration tech. In the neutral case, clean and efficient energy tech. In the worst case, climate-catastrophe coping tech.
Happy to discuss and help you where possible. How can I reach out to you? Or maybe you can send me a 'hi' at harsha.xg{at}gmail.com?
While I understand the importance of climate awareness and community building, I'm afraid that isn't where I shine.
Anything more technical, maybe in the fields of camera-traps, drone deployment, remote-sensing, satellite imagery, GIS, data analysis/viz etc may be a better fit. Trying to look for a way to get started in these.
Thanks again for taking the time to reply. Really appreciate it!
As for the copy, you're right. We need a lot of improvement on targeting. We thought LSPs and marketers are going to be main customers. Have to refine there. Our first idea was to do an API but we deferred it as the output still needs a pair of human eyes to approve.
Thank you for the feedback and wishes. Really appreciate it!
This was also the original problem we set out to solve. Thank you for letting us know that more people share the same problem. Unfortunately we haven't yet gotten it into the hands of the marketing agencies.
If you are ok with it, can you send us a 'hello' mail at team{at}imgtranslate.com so that we can discuss a little more with you.
Cheers
OCR -> Text property identification -> Inpainting -> Region averaging -> Translation -> Fitting
Please have the LSP try it out and see if offers some value.
As far as inpainting goes, we should soon be switching to GANs which produce structurally more accurate and less blurry results.
The problem with photoshop is that - You need a lof of fonts installed. - It doesn't always render the non-latin fonts correctly - It's expensive
We are little particular about what to translate and how to faithfully reproduce it. That's the main differentiation.
On design aspects, we wrote a short post here: http://blog.taglev.com/design-philosophy-behind-taglev/
In our use case, we've mostly had to deal with handwritten text and that's where none of them really did well. Your next best bet would be to use HoG(Histogram of oriented gradients) along with SVMs. OpenCV has really good implementations of both.
Even then, we've had to write extra heuristics to disambiguate between 2 and z and s and 5 etc. That was too much work and a lot of if-else. We're currently putting in our efforts on CNNs(Convolutional Neural Networks). As a start, you can look at Torch or Caffe.