Efficient Similarity Search in Multi-attribute Time Series Databases
Sang‐Jun Lee · The KIPS Transactions PartD · 2007
시계열에 대한 색인 및 검색 연구는 하나의 속성으로 구성된 시계열에 대하여 주로 수행되어 왔다. 그러나 음악, 비디오 등의 멀티미디어 데이타베이스는 다중속성 시계열 데이타베이스에서 유사 검색을 다룰 수 있어야 한다. 기존의 다중속성 시계열 데이타베이스에 대한 연구는 두 다중속성 시퀀스간의 유사도로 속성 간의 거리의 누적을 사용하고 있기에, 개별적인 속성 시퀀스에 대한 정보를 상실하게 된다. 본 연구에서는 이러한 문제를 해결하기 위해 속성 시퀀스 측면에서 다중속성 시계열 데이타베이스의 유사검색 기법을 제안한다. 제안된 기법은 검색 공간을 효율적으로 줄일 수 있으며, 착오 누락이 없음을 보장한다. 또한 실험을 통해 제안된 기법의 성능 향상을 확인하였다. Most of previous work on indexing and searching time series focused on the similarity matching and retrieval of one-attribute time series. However, multimedia databases such as music, video need to handle the similarity search in multi-attribute time series. The limitation of the current similarity models for multi-attribute sequences is that there is no consideration for attributes' sequences. The multi-attribute sequences are composed of several attributes' sequences. Since the users may want to find the similar patterns considering attributes's sequences, it is more appropriate to consider the similarity between two multi-attribute sequences in the viewpoint of attributes' sequences. In this paper, we propose the similarity search method based on attributes's sequences in multi-attribute time series databases. The proposed method can efficiently reduce the search space and guarantees no false dismissals. In addition, we give preliminary experimental results to show the effectiveness of the proposed method.