Running Fast SQL on DynamoDB Tables
rockset.com
rockset.com
It's effectively the same but technically quite different.
I ask because I have looked and haven't found one that I could follow without having to dedicate several hours to it.
To the point where I want to move off of dynamodb to aurora because I can't believe how hard it is to get simple insights like how many rows exist out of dynamo (I inherited the dynamo choice would not have used it myself)
Should take about an hour with testing to get a Pyspark script together to read in a DynamoDB table and write it out to S3.
https://docs.aws.amazon.com/glue/latest/dg/aws-glue-programm...
You’ll then need to crawl the S3 data to add it to your Glue catalog and then you can query it with Athena.
We switched to a scheduled Fargate task to dump data from DynamoDB into S3 as parquet files. It's really reliable, costs us ~$4/month and completely configurable.
https://aws.amazon.com/about-aws/whats-new/2018/07/aws-glue-...
Managing a database can be pretty expensive and time consuming.