Space-efficient construction of optimal prefix codes

Proceedings DCC '95 Data Compression Conference Pub Date : 1995-03-28 DOI:10.1109/DCC.1995.515509

Alistair Moffat, A. Turpin, J. Katajainen

引用次数: 20

Abstract

Shows that the use of the lazy list processing technique from the world of functional languages allows, under certain conditions, the package-merge algorithm to be executed in much less space than is indicated by the O(nL) space worst-case bound. For example, the revised implementation generates a 32-bit limited code for the TREC distribution within 15 Mb of memory. It is also shown how a second observation-that in large-alphabet situations it is often the case that there are many symbols with the same frequency-can be exploited to further reduce the space required, for both unlimited and length-limited coding. This second improvement allows calculation of an optimal length-limited code for the TREC word distribution in under 8 Mb of memory; and calculation of an unrestricted Huffman code in under 1 Mb of memory.

查看原文本刊更多论文

最优前缀码的空间高效构造

显示了使用函数式语言世界中的延迟列表处理技术，在某些条件下，包合并算法可以在比O(nL)空间最坏情况界所指示的更小的空间中执行。例如，修订后的实现在15 Mb内存内为TREC发行版生成32位限制代码。还展示了如何利用第二个观察结果——在大字母的情况下，通常会有许多具有相同频率的符号——来进一步减少无限编码和长度限制编码所需的空间。第二个改进允许在小于8mb的内存中计算TREC字分布的最佳长度限制代码;以及在1mb内存下计算一个不受限制的霍夫曼码。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Proceedings DCC '95 Data Compression Conference

自引率

0.00%

发文量