Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Cassandra was used at Twitter[0] to store quite a lot of time series data.

A typical production instance of the time series database is based on four distinct Cassandra clusters, each responsible for a different dimension (real-time, historical, aggregate, index) due to different performance constraints. These clusters are amongst the largest Cassandra clusters deployed in production today and account for over 500 million individual metric writes per minute. Archival data is stored at a lower resolution for trending and long term analysis, whereas higher resolution data is periodically expired.

[0]: https://blog.twitter.com/2013/observability-at-twitter



Believe they've moved to Manhattan, their own custom datastore:

https://blog.twitter.com/2014/manhattan-our-real-time-multi-...




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: