Answering multiple queries in compressed texts

  • Bin Wang*
  • , Minghe Yu
  • , Xiaochun Yang
  • , Guoren Wang
  • *Corresponding author for this work

Research output: Contribution to conferencePaperpeer-review

Abstract

With the exponential increment of data, compression technology becomes an important tool in the field of data management, especially in text management. An increasing pressing challenge is how to efficiently query these massive amounts of sequence data in their compressed format. In this paper we study the problem of answering subsequence-search queries on LZ78 format of texts. We propose the concept of conditional common sub strings of queries to improve query performance. We present a techniques to find minimal conditional common sub strings in compressed text and a local uncompressing technique to verify and locate positions of answers in text. Finally, the experimental results over real data demonstrate the efficiency of our algorithm.

Original languageEnglish
Pages61-66
Number of pages6
DOIs
Publication statusPublished - 2012
Externally publishedYes
Event9th Web Information Systems and Applications Conference, WISA 2012 - Haikou, Hainan, China
Duration: 16 Nov 201218 Nov 2012

Conference

Conference9th Web Information Systems and Applications Conference, WISA 2012
Country/TerritoryChina
CityHaikou, Hainan
Period16/11/1218/11/12

Keywords

  • common substring
  • multiple similar queries
  • string matching

Fingerprint

Dive into the research topics of 'Answering multiple queries in compressed texts'. Together they form a unique fingerprint.

Cite this