Indexed metadata

Strings with Maximally Many Distinct Subsequences and Substrings

Abraham Flaxman, Aram W. Harrow, Gregory B. Sorkin

Source record

Source: Crossref

Published: Jan 5, 2004

DOI: 10.37236/1761

Open original source ↗

Source abstract

A natural problem in extremal combinatorics is to maximize the number of distinct subsequences for any length-nn string over a finite alphabet Σ\Sigma; this value grows exponentially, but slower than 2n2^n. We use the probabilistic method to determine the maximizing string, which is a cyclically repeating string. The number of distinct subsequences is exactly enumerated by a generating function, from which we also derive asymptotic estimates. For the alphabet Σ={1,2}\Sigma=\{1,2\}, (1,2,1,2,)\,(1,2,1,2,\dots) has the maximum number of distinct subsequences, namely Fib(n+3)1((1+5)/2)n+3 ⁣/5{\rm Fib}(n+3)-1 \sim \left((1+\sqrt5)/2\right)^{n+3} \! / \sqrt{5}. We also consider the same problem with substrings in lieu of subsequences. Here, we show that an appropriately truncated de Bruijn word attains the maximum. For both problems, we compare the performance of random strings with that of the optimal ones.

Evidence graph

No public relationships recorded yet.

Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.

Strings with Maximally Many Distinct Subsequences and Substrings — Mathematical Frontier Network