Reconsidering the significance of genomic word frequency

dc.creatorCsűrös, Miklós
dc.creatorNoé, Laurent
dc.creatorKucherov, Gregory
dc.date2006-09-14
dc.date.accessioned2026-07-07T07:26:22Z
dc.date.available2026-07-07T07:26:22Z
dc.descriptionWe propose that the distribution of DNA words in genomic sequences can be primarily characterized by a double Pareto-lognormal distribution, which explains lognormal and power-law features found across all known genomes. Such a distribution may be the result of completely random sequence evolution by duplication processes. The parametrization of genomic word frequencies allows for an assessment of significance for frequent or rare sequence motifs.
dc.identifierhttps://arxiv.org/abs/q-bio/0609022
dc.identifierhttp://arxiv.org/abs/q-bio/0609022
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/117048
dc.subjectGenomics
dc.titleReconsidering the significance of genomic word frequency
dc.typetext

Files

Collections