UM  > 科技學院  > 電腦及資訊科學系
Discovering longestlasting correlation in sequence databases
Li Y.1; Leong Hou U.1; Yiu M.L.2; Gong Z.1
2013
Source PublicationProceedings of the VLDB Endowment
ISSN21508097
Volume6Issue:14Pages:1666
Abstract

Most existing work on sequence databases use correlation (e.g., Euclidean distance and Pearson correlation) as a core function forvarious analytical tasks. Typically, it requires users to set a length for the similarity queries. However, there is no steady way to define the proper length on different application needs. In this work we focus on discovering longest-lasting highly correlated subsequences in sequence databases, which is particularly useful in helping those analyses without prior knowledge about the query length. Surprisingly, there has been limited work on this problem. A baseline solution is to calculate the correlations for every possible subsequence combination. Obviously, the brute force solution is not scalable for large datasets. In this work we study a space-constrained index that gives a tight correlation bound for subsequences of similar length and offset by intra-object grouping and inter-object grouping techniques. To the best of our knowledge, this is the first index to support normalized distance metric of arbitrary length subsequences.Extensive experimental evaluation on both real and synthetic sequence datasets verifies the efficiency and effectiveness of our proposed methods. © 2013 VLDB Endowment.

DOI10.14778/2556549.2556552
URLView the original
Indexed BySCI
Language英语
The Source to ArticleScopus
Fulltext Access
Citation statistics
Document TypeJournal article
CollectionDEPARTMENT OF COMPUTER AND INFORMATION SCIENCE
Affiliation1.Department of Computer and Information Science, University of Macau, Macau
2.Department of Computing, Hong Kong Polytechnic University, Hung Hom, Kowloon, Hong Kong
First Author AffilicationUniversity of Macau
Recommended Citation
GB/T 7714
Li Y.,Leong Hou U.,Yiu M.L.,et al. Discovering longestlasting correlation in sequence databases[J]. Proceedings of the VLDB Endowment,2013,6(14):1666.
APA Li Y.,Leong Hou U.,Yiu M.L.,&Gong Z..(2013).Discovering longestlasting correlation in sequence databases.Proceedings of the VLDB Endowment,6(14),1666.
MLA Li Y.,et al."Discovering longestlasting correlation in sequence databases".Proceedings of the VLDB Endowment 6.14(2013):1666.
Related Services
Recommend this item
Bookmark
Usage statistics
Export to Endnote
Google Scholar
Similar articles in Google Scholar
[Li Y.]'s Articles
[Leong Hou U.]'s Articles
[Yiu M.L.]'s Articles
Baidu academic
Similar articles in Baidu academic
[Li Y.]'s Articles
[Leong Hou U.]'s Articles
[Yiu M.L.]'s Articles
Bing Scholar
Similar articles in Bing Scholar
[Li Y.]'s Articles
[Leong Hou U.]'s Articles
[Yiu M.L.]'s Articles
Terms of Use
No data!
Social Bookmark/Share
All comments (0)
No comment.
 

Items in the repository are protected by copyright, with all rights reserved, unless otherwise indicated.