Representing Text Chunks

dc.creatorSang, Erik F. Tjong Kim
dc.creatorVeenstra, Jorn
dc.date1999-07-06
dc.date.accessioned2026-07-07T03:24:12Z
dc.date.available2026-07-07T03:24:12Z
dc.descriptionDividing sentences in chunks of words is a useful preprocessing step for parsing, information extraction and information retrieval. (Ramshaw and Marcus, 1995) have introduced a "convenient" data representation for chunking by converting it to a tagging task. In this paper we will examine seven different data representations for the problem of recognizing noun phrase chunks. We will show that the the data representation choice has a minor influence on chunking performance. However, equipped with the most suitable data representation, our memory-based learning chunker was able to improve the best published chunking results for a standard data set.
dc.description7 pages
dc.identifierhttps://arxiv.org/abs/cs/9907006
dc.identifierhttp://arxiv.org/abs/cs/9907006
dc.identifierEACL'99, Bergen
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/33235
dc.subjectComputation and Language
dc.subjectI.2.7
dc.titleRepresenting Text Chunks
dc.typetext

Files

Collections