Similar sequence matching supporting variable-length and variable-tolerance continuous queries on time-series data stream

Citations

WEB OF SCIENCE

15
Citations

SCOPUS

21

초록

We propose a new similar sequence matching method that efficiently supports variable-length and variable-tolerance continuous query sequences on time-series data stream. Earlier methods do not support variable lengths or variable tolerances adequately for continuous query sequences if there are too many query sequences registered to handle in main memory. To support variable-length query sequences, we use the window construction mechanism that divides long sequences into smaller windows for indexing and searching the sequences. To support variable-tolerance query sequences, we present a new notion of intervaled sequences whose individual entries are an interval of real numbers rather than a real number itself. We also propose a new similar sequence matching method based on these notions, and then, formally prove correctness of the method. In addition, we show that our method has the prematching characteristic, which finds future candidates of similar sequences in advance. Experimental results show that our method outperforms the naive one by 2.6-102.1 times and the existing methods in the literature by 1.4-9.8 times over the entire ranges of parameters tested when the query selectivities are low (< 32%), which are practically useful in large database applications. (c) 2007 Published by Elsevier Inc.

키워드

similar sequence matchingtime-series datadata streamscontinuous queries
제목
Similar sequence matching supporting variable-length and variable-tolerance continuous queries on time-series data stream
저자
Lim, Hyo-SangWhang, Kyu-YoungMoon, Yang-Sae
DOI
10.1016/j.ins.2007.10.026
발행일
2008-03-15
유형
Article
저널명
Information Sciences
178
6
페이지
1461 ~ 1478