337 karma · joined April 30, 2015
This is a well-known term for everyone working with larger datasets...
I‘m running it on AWS Lambda functions with some success.
I‘m using a (older) v1.29.1 dev version with https://sql-workbench.com w/o any bigger issues.
You can either drag & drop data, or use remote data sources via https
A recent project is https://shrink.video, which is using the WASM version of ffmpeg to shrink or convert video in the user's browser itself, for privacy and similar reasons mentioned before.
It‘s a SQL Workbench in the browser, based on DuckDB WASM. You can query remote and local datasources, such as CSV, JSON or Parquet files.
You can also visualize the results, and share the queries via URL. Let me know what you think!
With 1TB free traffic from the CDN, and pretty small costs for S3 and Lambda@Edge, it's probably even cheaper to self-host I guess. Even lower costs would be possible by entirely using CloudFlare services (CDN, R2)...
There's also https://www.awsiamdata.com/ which analyzes the AWS IAM data and also contains a changelog. Maybe this helps some people.
You can also query the data from your browser: https://tobilg.com/chat-with-a-duck#heading-explore-aws-iam-...
See https://tobilg.com/using-duckdb-wasm-for-in-browser-data-eng...
Otherwise, I‘m very open to feedback, just create an issue in the linked GH Repo. Thanks!