I have a Fujitsu ScanSnap which is one of those feed-through scanners. I have it hooked up to a Raspberry Pi which listens for the button press on the scanner. You press the button, the paper feeds through the scanner and once it has finished the scan a script runs to collate everything into a PDF and drops the result onto a Samba share that's running on the box where paperless-ngx is.
It's pretty neat and feels seamless. The worst part was dealing with SANE and finding linux drivers for my scanner.
What does listen for button press mean? and how?
This works fine on say, a Mac, with the official Fujitsu ScanSnap software, and I'm guessing _that_ supports saving to a samba share, but I wanted a solution that's
1. completely headless, i.e. no desktop machine required and experience needs to be friction free as the headless part means the only way to interact with the scanning function is to press 1 button
2. linux compatible, as I wanted to connect it to a Pi. I had to dig for the drivers, Fujitsu didn't have the right ones for my model on their website!
I couldn't find any official software from Fujitsu, but I found the drivers eventually, so ended up coming up with connecting the scanner to the Pi over USB and glueing the bits together to drop the PDFs onto the samba share
The button is located on the scanner, and I run "scanbd" [1] to listen for the button press, this is what coordinates the scan function (feeding the paper through) and then post-scan -> running a script to collate + create PDFs
https://chrisschuld.com/2020/01/network-scanner-with-scansna...
I've switched my hand-crafted scripts recently to use scanpdf[1] which seems to give better results (once I tweaked it to be a little less eager to downconvert to B+W). I experimented with using OpenCV models for cropping and straightening (based on examples in a stackoverflow thread at [2]) but I found results were worse than scanpdf so far.
1. http://badge.fury.io/py/scanpdf 2. https://stackoverflow.com/questions/28935983/preprocessing-i...
I scan into a Samba share that paperless-ngx picks up automatically, OCRs, tags, and deletes.
A web application is pretty cross platform too, at this point.
Plus I can get to them on my phones with less trouble than a share.
How does it handle when you have digital documents you want to store (a la google drive or similar)?