Statistical composition of high-scoring segments from molecular sequences

From MaRDI portal





New probabilistic formulas that provide a benchmark for discerning distributional properties of various data statistics or letter sequences (e.g., DNA) are presented. These include the asymptotical extremal distribution of high aggregate segment scores and the limiting letter composition of high-scoring segments, as well as a number of associated conditional Gaussian central limit laws. These formulas are derived with respect to a general scoring scheme with i.i.d. letter values.











This page was built for publication: Statistical composition of high-scoring segments from molecular sequences

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q922997)