首页 | 本学科首页   官方微博 | 高级检索  
     


Sublinear Algorithms for Approximating String Compressibility
Authors:Sofya Raskhodnikova  Dana Ron  Ronitt Rubinfeld  Adam Smith
Affiliation:1. Pennsylvania State University, University Park, PA, USA
2. Tel Aviv University, Tel Aviv, Israel
3. MIT, Cambridge, MA, USA
Abstract:We raise the question of approximating the compressibility of a string with respect to a fixed compression scheme, in sublinear time. We study this question in detail for two popular lossless compression schemes: run-length encoding (RLE) and a variant of Lempel-Ziv (LZ77), and present sublinear algorithms for approximating compressibility with respect to both schemes. We also give several lower bounds that show that our algorithms for both schemes cannot be improved significantly. Our investigation of LZ77 yields results whose interest goes beyond the initial questions we set out to study. In particular, we prove combinatorial structural lemmas that relate the compressibility of a string with respect to LZ77 to the number of distinct short substrings contained in it (its ?th subword complexity , for small ?). In addition, we show that approximating the compressibility with respect to LZ77 is related to approximating the support size of a distribution.
Keywords:
本文献已被 SpringerLink 等数据库收录!
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号