How to Partitioning Data for Linear Scalability in Geospatial Queries?
The other potential solution is to overlap data so a node contains the tiles along its edges from the next and previous nodes as well. Not 100% sure how to handle this, what the best technology is etc.
Any recommendations welcome. I'm probably looking at the problem wrong - eg that a partition key in a columnar database query (eg cassandra) may be the floored lat & long integers getting a column range of the lesser significant digits. But maybe there is another way of looking at the data/problem space?