On the Limiting Distribution of Lempel-Ziv'78 Redundancy for Memoryless Sources

被引:3
|
作者
Jacquet, Philippe [1 ]
Szpankowski, Wojciech [2 ,3 ]
机构
[1] Alcatel Lucent Bell Labs, F-91620 Nozay, France
[2] Purdue Univ, Dept Comp Sci, W Lafayette, IN 47907 USA
[3] Gdansk Univ Technol, Fac Elect Telecommun & Informat, PL-80233 Gdansk, Poland
基金
美国国家科学基金会;
关键词
Lempel-Zv; 78; redundancy; digital search trees; analytic information theory; ZIV PARSING SCHEME; PROBABILITY; SEQUENCES;
D O I
10.1109/TIT.2014.2358679
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
We study the Lempel-Ziv'78 algorithm and show that its (normalized) redundancy rate tends to a Gaussian distribution for memoryless sources. We accomplish it by extending findings from our 1995 paper, in particular, by presenting a new simplified proof of the central limit theorem (CLT) for the number of phrases in the LZ'78 algorithm. We first analyze the asymptotic behavior of the total path length in the associated digital search tree built from independent sequences. Then, a renewal theory type argument yields CLT for LZ'78 scheme. Here, we extend our analysis of LZ'78 algorithm to present new results on the convergence of moments, moderate and large deviations, and CLT for the (normalized) redundancy. In particular, we confirm that the average redundancy rate decays as 1/log n, and we find that the variance is of order 1/n, where n is the length of the text.
引用
收藏
页码:6917 / 6930
页数:14
相关论文
共 50 条
  • [41] ESTIMATION OF THE ENTROPY BY THE LEMPEL-ZIV METHOD
    HANSEL, G
    LECTURE NOTES IN COMPUTER SCIENCE, 1989, 377 : 51 - 65
  • [42] Automata on Lempel-Ziv compressed strings
    Leiss, H
    de Rougemont, M
    COMPUTER SCIENCE LOGIC, PROCEEDINGS, 2003, 2803 : 384 - 396
  • [43] Engineering Practical Lempel-Ziv Tries
    Arroyuelo D.
    Cánovas R.
    Fischer J.
    Köppl D.
    Löbel M.
    Navarro G.
    Raman R.
    1600, Association for Computing Machinery (26):
  • [44] Lempel-Ziv compression of highly structured documents
    Adiego, Joaquin
    Navarro, Gonzalo
    de la Fuente, Pablo
    JOURNAL OF THE AMERICAN SOCIETY FOR INFORMATION SCIENCE AND TECHNOLOGY, 2007, 58 (04): : 461 - 478
  • [45] Lempel-Ziv index for q-grams
    Karkkainen, J
    Sutinen, E
    ALGORITHMICA, 1998, 21 (01) : 137 - 154
  • [46] The Lempel-Ziv complexity of fixed points of morphisms
    Constantinescu, Sorin
    Ilie, Lucian
    MATHEMATICAL FOUNDATIONS OF COMPUTER SCIENCE 2006, PROCEEDINGS, 2006, 4162 : 280 - 291
  • [47] Lossy Lempel-Ziv algorithm for large alphabet sources and applications to image compression
    Finamore, WA
    Leister, MD
    INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, PROCEEDINGS - VOL I, 1996, : 225 - 228
  • [48] Practical fixed length Lempel-Ziv coding
    Klein, Shmuel T.
    Shapira, Dana
    DISCRETE APPLIED MATHEMATICS, 2014, 163 : 326 - 333
  • [49] ON THE BIT-COMPLEXITY OF LEMPEL-ZIV COMPRESSION
    Ferragina, Paolo
    Nitto, Igor
    Venturini, Rossano
    SIAM JOURNAL ON COMPUTING, 2013, 42 (04) : 1521 - 1541
  • [50] The Lempel-Ziv complexity in infinite ergodic systems
    Shinkai, Soya
    Aizawa, Yoji
    JOURNAL OF THE KOREAN PHYSICAL SOCIETY, 2007, 50 (01) : 261 - 266