Clustered Language Models with Context-Equivalent States

dc.creatorUeberla, J. P.
dc.creatorGransden, I. R.
dc.date1996-06-04
dc.date.accessioned2026-07-07T09:10:21Z
dc.date.available2026-07-07T09:10:21Z
dc.descriptionIn this paper, a hierarchical context definition is added to an existing clustering algorithm in order to increase its robustness. The resulting algorithm, which clusters contexts and events separately, is used to experiment with different ways of defining the context a language model takes into account. The contexts range from standard bigram and trigram contexts to part of speech five-grams. Although none of the models can compete directly with a backoff trigram, they give up to 9\% improvement in perplexity when interpolated with a trigram. Moreover, the modified version of the algorithm leads to a performance increase over the original version of up to 12\%.
dc.description3 pages, latex
dc.identifierhttps://arxiv.org/abs/cmp-lg/9606002
dc.identifierhttp://arxiv.org/abs/cmp-lg/9606002
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/151325
dc.subjectComputation and Language
dc.titleClustered Language Models with Context-Equivalent States
dc.typetext

Files

Collections