Yhat Sciencebox
blog.yhathq.com
blog.yhathq.com
My current setup is to use a workhorse server where I set up a samba share for my home folder (so I can edit scripts remotely from my notebook without the git-push-pull dance) and a remote iPython notebook. Both config files are write once and forget and I get 90% of their functionality built in (without the web GUI). Would love to learn about other people's setup though, I am sure mine can be improved.
Neat idea, I just find it difficult to justify a $1000/mo per-machine cost, especially in academia.
However, the incremental improvement of renting an EC2 instance (even with 8 cores and 65gb) pales in comparison to using a distributed data processing approach with medium to large datasets. Any plans for supporting Hadoop/Spark + Amazon EMR in the future?
Otherwise -- it's definitely something we'd consider using; I end up doing 99% of my work on my laptop because AWS setup fixed cost time is 30minutes plus every time.
And perhaps updating the COBOL scripts from 1970-something that I inherited