Conferences >2013 Data Compression Confere...

Faster Compact Top-k Document Retrieval

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

An optimal index solving top-k document retrieval [Navarro and Nekrich, SODA'12] takes O(m+k) time for a pattern of length m, but its space is at least 80n bytes for a co...Show More

Metadata

Abstract:

An optimal index solving top-k document retrieval [Navarro and Nekrich, SODA'12] takes O(m+k) time for a pattern of length m, but its space is at least 80n bytes for a collection of n symbols. We reduce it to 1.5n-3n bytes, with O(m + (k+log log n)log log n) time, on typical texts. The index is up to 25 times faster than the best previous compressed solutions, and requires at most 5% more space in practice (and in some cases as little as one half). Apart from replacing classical by compressed data structures, our main idea is to replace suffix tree sampling by frequency thresholding to achieve compression.

Published in: 2013 Data Compression Conference

Date of Conference: 20-22 March 2013

Date Added to IEEE Xplore: 20 June 2013

ISBN Information:

Print ISSN: 1068-0314

DOI: 10.1109/DCC.2013.43

Conference Location: Snowbird, UT, USA

Contents

References is not available for this document.

Faster Compact Top-k Document Retrieval

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?

Faster Compact Top-k Document Retrieval

Alerts

Abstract:

Metadata

Abstract:

Authors

Figures

References

Citations

Keywords

Metrics

Footnotes

References

IEEE Account

Purchase Details

Profile Information

Need Help?