RethinkDB screencast - from queries to sharding under 15 minutes
rethinkdb.com
rethinkdb.com
Anybody know of any demos of RethinkDB handling, say, 100gb of data? And running decent queries on it?
The underlying storage engine was tested on commodity systems and super-duper enterprisy storage systems, and can do hundreds of thousands of ops/second on tens of terabytes of data (that required pretty beefy setups, though). When we added clustering on top of the storage engine, we avoided thinking of performance too much (in the interest of shipping), so everything slowed down significantly. Here's our (rough) roadmap:
- New protocol buffer API and some more checklist features (1.4)
- Secondary indexes, huge ReQL improvements (1.5)
- Performance and scalability (1.6)
We'll be doing scalability and performance demos that I hope will be really impressive, but it'll take ~4 months to get there.Can you guys add a "rough roadmap overview" page to the docs, so we could have a general idea what is the status?
I like the way the RubyMine does it:
http://confluence.jetbrains.net/display/RUBYDEV/Development+...
"RethinkDB is a great choice if you .... are planning to run anywhere from a single node to sixteen node clusters."
With a sharded master-slave setup with one slave each, this leaves us with a total of 8 shards. This is enough for most use cases, but is there a reason it is limited to 16 nodes?
16 is the largest number of nodes we've done sufficiently rigorous tests on to be sure that things go smoothly. So that's the highest number we're comfortable citing on our site. It's a conservative estimate though so you should be fine straying past it.
We know of a few things which becoming scalability concerns with a large numbers of machines but we're talking close to 100 machines. These will hopefully be addressed soon.
In the website it stated as 1.3.2 (which imply production ready) but I think I saw some comments a month ago from you that it's not fit for production use yet.
What about secondary indexes?
Are the machines in the screencast are very weak? a simple query (get 2 rows of the dota table) running ~100ms is really slow- is it because you're using the web interface?
RethinkDB seems cool and I really want to try it in my next pet project :)
Not yet. We'll bump the release to 2.0 when it's ready for production.
> What about secondary indexes?
They're coming -- see https://github.com/rethinkdb/rethinkdb/issues/88
> Are the machines in the screencast are very weak?
No, the 100ms roundtrip includes the HTTP request over our admittedly very unsophisticated WiFi network.
Hope this helps!
- p.s. love the demo video looks awesome how easy it is to use and I like the query language you created looks nice too.
There are still many improvements we can make to the core system and an enormous number of people are already interested in it, so we decided to satisfy them first. We might add a SQL front-end at some point, but it's not very high on the priority list now.