Text Segmentation Based on Similarity between Words

dc.creatorKozima, Hideki
dc.date1996-01-23
dc.date.accessioned2026-07-07T09:10:05Z
dc.date.available2026-07-07T09:10:05Z
dc.descriptionThis paper proposes a new indicator of text structure, called the lexical cohesion profile (LCP), which locates segment boundaries in a text. A text segment is a coherent scene; the words in a segment are linked together via lexical cohesion relations. LCP records mutual similarity of words in a sequence of text. The similarity of words, which represents their cohesiveness, is computed using a semantic network. Comparison with the text segments marked by a number of subjects shows that LCP closely correlates with the human judgments. LCP may provide valuable information for resolving anaphora and ellipsis.
dc.description3 pages, uufiles (paper.tex, acl.sty, bezier.sty)
dc.identifierhttps://arxiv.org/abs/cmp-lg/9601005
dc.identifierhttp://arxiv.org/abs/cmp-lg/9601005
dc.identifierProceedings of ACL-93 (Ohio), pp.286-288, 1993.
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/151258
dc.subjectComputation and Language
dc.titleText Segmentation Based on Similarity between Words
dc.typetext

Files

Collections