Cypress: Managing Massive Time Series Streams with MultiScale Compressed Trickles
- Galen Reeves ,
- Jie Liu ,
- Suman Nath ,
- Feng Zhao
MSR-TR-2009-79 |
We present Cypress, a novel framework to archive and query massive time series streams such as those generated by sensor networks, data centers, and scientific computing. Cypress applies multi-scale analysis to decompose time series and to obtain sparse representations in various domains (e.g. frequency domain and time domain). Relying on the sparsity, the time series streams can be archived with reduced storage space. We then show that many statistical queries such as trend, histogram and correlations can be answered directly from compressed data rather than from reconstructed raw data. Our evaluation with server utilization data collected from real data centers shows significant benefit of our framework.