Paperless-ngx: scan, index and archive all your physical documents
github.com
github.com
I suppose I could convert "finished" Google docs to PDF and save them in paperless, but it just seems like these systems will always be disconnected in some way.
I've even experimenting a generic usage of Zotero, does not work much, ideally these days we need to manage files NOT in a hierarchy but in a graph, automatically managed, annotating files with links in notes, being able to search titles, links, notes all together.
Zim with attachments for non-teaches it's still limited, too tied to the underlying file system, Zotero and Paperless are way too mechanic and Paperless do not allow note, a separate Dokuwiki with links to Paperless stored docs it's simply way too much overhead...
Long story short it's remarkable the automatic OCR (ocrmypdf), auto-classification, metadata automation etc but it's still not "the universal solution" IMO...
I've been itching to give paperless-ngx a shot because I just love it but ldap hasn't yet ended up in the docs but the pull request was merged https://github.com/paperless-ngx/paperless-ngx/pull/5190.
Regardless, I just love how this project just keeps coming back to life
LDAP is such a pain in the ass to integrate with, and it seems like most things are going OIDC these days.
I've been avoiding LDAP like the plague. I think MS is moving away from self-hosted AD, and LDAP really loses its luster for most folks when the self hosted options are something like OpenLDAP.
And the parent makes a good point that OIDC/OAuth does not give group membership.
Kerberos, yes, but LDAP no.
What are your pain points integrating with LDAP? It is pretty simple.
LDAP is a pain because you have to expose/support a lot of knobs for integration (bind vs anonymous, secure vs unsecure, group format, root DNs, etc.). OIDC is (in theory) a lot simpler for the most part as the bare minimum is discovery URL, client ID, and client secret.
> The easiest way to deploy paperless is docker compose
Ok, that's a first red flag.
I got a Brother ADS-4300N to use with p-ngx, works very well and also is way faster than the usual document scanners on MF printers (duplex is done in one pass, for example)...
set up was straightforward and the functionality is great
And "docker compose up" is the easiest way to deploy things these days in general. That's got nothing to do with this software specifically.
You don't want to use paperless-ngx for editable stuff really. You want to use it for stuff like bills, invoices, and business records.
Once it's in paperless, it's searchable and you don't have to worry about where it is. As long as the scan is good it will grab the OCR and then you can search for things like account number. My uncle basically scans everything bill related into his instance and then shreds the paper.
You can also tag documents and search by tag. Also since it's a web app if you can do the self-hosted thing it works well on the phone.
I wrote an iOS [1] app to connect to you instance and it’s open source [2].
[1] https://apps.apple.com/de/app/swift-paperless/id6448698521
> Documents are saved as PDF/A format which is designed for long term storage…[snip]
Can someone please tell me what attributes make a given file format more suitable for long term storage over another?