Movatterモバイル変換

Smallest grammar problem

From Wikipedia, the free encyclopedia

Indata compression and the theory offormal languages, thesmallest grammar problem is the problem of finding the smallestcontext-free grammar that generates a givenstring of characters (but no other string). The size of a grammar is defined by some authors as the number of symbols on the right side of the production rules.^[1]Others also add the number of rules to that.^[2] A grammar that generates only a single string, as required for the solution to this problem, is called astraight-line grammar.^[3]

Everybinary string of length $n {\displaystyle n}$ has a grammar of length $O(n/\log n)$ , as expressed usingbig O notation.^[3] For binaryde Bruijn sequences, no better length is possible.^[4]

The (decision version of the) smallest grammar problem isNP-complete.^[1]It can be approximated inpolynomial time to within a logarithmicapproximation ratio; more precisely, the ratio is $O(\log {\tfrac {n}{g}})$ where $n {\displaystyle n}$ is the length of the given string and $g {\displaystyle g}$ is the size of its smallest grammar. It is hard to approximate to within a constant approximation ratio. An improvement of the approximation ratio to $o(\log n/\log \log n)$ would also improve certain algorithms for approximateaddition chains.^[5]

References

[edit]

^^a ^bCharikar, Moses; Lehman, Eric; Liu, Ding; Panigrahy, Rina; Prabhakaran, Manoj; Sahai, Amit; Shelat, Abhi (2005)."The Smallest Grammar Problem".IEEE Transactions on Information Theory.51 (7):2554–2576.CiteSeerX 10.1.1.185.2130.doi:10.1109/TIT.2005.850116.S2CID 6900082.Zbl 1296.68086.
^Florian Benz and Timo Kötzing, “An effective heuristic for the smallest grammar problem,” Proceedings of the fifteenth annual conference on Genetic and evolutionary computation conference - GECCO ’13, 2013.ISBN 978-1-4503-1963-8doi:10.1145/2463372.2463441
^^a ^bLohrey, Markus (2012)."Algorithmics on SLP-compressed strings: A survey"(PDF).Groups Complexity Cryptology.4 (2):241–299.doi:10.1515/GCC-2012-0016.
^Domaratzki, Michael; Pighizzini, Giovanni; Shallit, Jeffrey (2002). "Simulating finite automata with context-free grammars".Information Processing Letters.84 (6):339–344.doi:10.1016/S0020-0190(02)00316-2.MR 1937222.
^Charikar, Moses; Lehman, Eric; Liu, Ding; Panigrahy, Rina; Prabhakaran, Manoj; Rasala, April; Sahai, Amit; Shelat, Abhi (2002)."Approximating the Smallest Grammar: Kolmogorov Complexity in Natural Models"(PDF).Proceedings of the thirty-fourth annual ACM symposium on theory of computing (STOC 2002), Montreal, Quebec, Canada, May 19–21, 2002. New York, NY: ACM Press. pp. 792–801.doi:10.1145/509907.510021.ISBN 978-1-581-13495-7.S2CID 282489.Zbl 1192.68397.

External links

[edit]

"CFG-Kolm-complexity is singleton sets with Lance and Bill".Computational Complexity. June 9, 2024.

Data compression methods

Lossless
type

Entropy	Adaptive coding Arithmetic Asymmetric numeral systems Golomb Huffman Adaptive Canonical Modified Range Shannon Shannon–Fano Shannon–Fano–Elias Tunstall Unary Universal Exp-Golomb Fibonacci Gamma Levenshtein
Dictionary	Byte-pair encoding Lempel–Ziv 842 LZ4 LZJB LZO LZRW LZSS LZW LZWL Snappy
Other	BWT CTW CM Delta Incremental DMC DPCM Grammar Re-Pair Sequitur LDCT MTF PAQ PPM RLE
Hybrid	LZ77 + Huffman Deflate LZX LZS LZ77 + ANS LZFSE LZ77 + Huffman + ANS Zstandard LZ77 + Huffman + context Brotli LZSS + Huffman LHA/LZH LZ77 + Range LZMA LZHAM RLE + BWT + MTF + Huffman bzip2

Lossy
type

Transform	Discrete cosine transform DCT MDCT DST FFT Wavelet Daubechies DWT SPIHT
Predictive	DPCM ADPCM LPC ACELP CELP LAR LSP WLPC Motion Compensation Estimation Vector Psychoacoustic

Audio

Concepts	Bit rate ABR CBR VBR Companding Convolution Dynamic range Latency Nyquist–Shannon theorem Sampling Silence compression Sound quality Speech coding Sub-band coding
Codec parts	A-law μ-law DPCM ADPCM DM FT FFT LPC ACELP CELP LAR LSP WLPC MDCT Psychoacoustic model

Image

Concepts	Chroma subsampling Coding tree unit Color space Compression artifact Image resolution Macroblock Pixel PSNR Quantization Standard test image Texture compression
Methods	Chain code DCT Deflate Fractal KLT LP RLE Wavelet Daubechies DWT EZW SPIHT

Video

Concepts	Bit rate ABR CBR VBR Display resolution Frame Frame rate Frame types Interlace Video characteristics Video quality
Codec parts	DCT DPCM Deblocking filter Lapped transform Motion Compensation Estimation Vector Wavelet Daubechies DWT