How do you handle the privacy of the scanned documents?
One level of privacy is the workspace level separation in Mongo. But, if there is customer interest, other setups are possible. E.g. the way Databricks handles privacy is by actually giving each account its own back end services - and scoping workspaces within an account.
That is a good possible model.
- data is never shared between customers
- data never gets used for training
- we also configure data retention policies to auto-purge after a time period