Convert paper-based notes to HTML content with Google Vision API
itnext.io
itnext.io
I wrote this app to scan the labels and read them out loud so my grandpa can manage his life.
https://www.helpmereadthis.com/
It’s open source, so feel free to steal code.
I like/prefer it a lot when you can simply put up a script, press 3 buttons and voila you have a fully working url. Even Google used to do that, I've built two libraries around Google's services, npm's `translate`[1] and `drive-db`[2], the former I no longer test on Google Translate even though it's supposed to be supported and the later it's using a more obscure plain JSON API to avoid Google Auth.
I can imagine one necessary enhancement (I would find it necessary anyway) would be to support crossing-through words with a line to "erase" them.
I like the local aspect, one project I have in mind is a hand writing to printed letters for note taking/tablets. Not a new concept I think one drive has it, but it would be neat to train it on your own writing.
Why not Zoidberg
[0] https://www.nist.gov/itl/products-and-services/emnist-datase...
Obviously the tech can always be improved, but I find this to be a secondary concern. Ideas can still be cool without worrying about implementation details. For instance, if the PoC works well enough, but you want better performance, then you can train a new OCR system specifically for it. You can break the project down into pieces, where each piece is an interesting challenge. Meanwhile getting 80% performance out of an existing solution allows you to build up a dataset to be used for improving it.
The title is funny to me, because I know people who deliberately take paper notes to keep them safe from Google.
I guess now they have an option should they change their views on privacy.
This seems like it could be nice to tinker with building something similar.
I prefer hand written notes but prefer to be able recall the notes digitally from a search.
And, it's getting old, but why does Google need to know any notes I write down. My desktop computer and phone seem powerful enough to do any OCR with the right software.
So they can target you with more ads.
MacOS:
https://apps.apple.com/us/app/image-text-ocr-photo-pdf-scan/...
iOS:
https://apps.apple.com/us/app/image-text-ocr-photo-scanner/i...
[0]: https://outline.com/KZqa5L
Edit:
[1]: Images on Medium can be resized by changing the expected width param, as in https://miro.medium.com/max/{width}/{filename} URI.
- Outline-generated link (width = 60 px): https://miro.medium.com/max/60/1*rNmb72I6rBUckbJuI6hRFw.png?...
- Resized to 1024px width: https://miro.medium.com/max/1024/1*rNmb72I6rBUckbJuI6hRFw.pn...
- incognito tab in chrome
- delete the medium or in this case itnext.io cookie
I don't want to have to use a Google account and to share my images and notes and text with Google and the US government.
Also - I agree with other commenters that it'll probably be a better idea to obtain the "unparsed" markup (or should I say - markdown?) rather than generating HTML.