A Regression-Based Temporal Pattern Mining Scheme for Data Streams [chapter]

Wei-Guang Teng, Ming-Syan Chen, Philip S. Yu
2003 Proceedings 2003 VLDB Conference  
We devise in this paper a regression-based algorithm, called algorithm FTP-DS (Frequent Temporal Patterns of Data Streams), to mine frequent temporal patterns for data streams. While providing a general framework of pattern frequency counting, algorithm FTP-DS has two major features, namely one data scan for online statistics collection and regressionbased compact pattern representation. To attain the feature of one data scan, the data segmentation and the pattern growth scenarios are explored
more » ... or the frequency counting purpose. Algorithm FTP-DS scans online transaction flows and generates candidate frequent patterns in real time. The second important feature of algorithm FTP-DS is on the regression-based compact pattern representation. Specifically, to meet the space constraint, we devise for pattern representation a compact ATF (standing for Accumulated Time and Frequency) form to aggregately comprise all the information required for regression analysis. In addition, we develop the techniques of the segmentation tuning and segment relaxation to enhance the functions of FTP-DS. With these features, algorithm FTP-DS is able to not only conduct mining with variable time intervals but also perform trend detection effectively. Synthetic data and a real dataset which contains net-work alarm logs from a major telecommunication company are utilized to verify the feasibility of algorithm FTP-DS.
doi:10.1016/b978-012722442-8/50017-3 dblp:conf/vldb/TengCY03 fatcat:vhafzvjx3ra7lgtkfzlrt5t5di