Automatic Extraction of Tagset Mappings from Parallel-Annotated Corpora

dc.creatorHughes, John
dc.creatorSouter, Clive
dc.creatorAtwell, Eric
dc.date1995-06-08
dc.date.accessioned2026-07-07T09:09:54Z
dc.date.available2026-07-07T09:09:54Z
dc.descriptionThis paper describes some of the recent work of project AMALGAM (automatic mapping among lexico-grammatical annotation models). We are investigating ways to map between the leading corpus annotation schemes in order to improve their resuability. Collation of all the included corpora into a single large annotated corpus will provide a more detailed language model to be developed for tasks such as speech and handwriting recognition. In particular, we focus here on a method of extracting mappings from corpora that have been annotated according to more than one annotation scheme.
dc.description8 pages, LaTeX, uses EACL95 style file: aclap.sty
dc.identifierhttps://arxiv.org/abs/cmp-lg/9506006
dc.identifierhttp://arxiv.org/abs/cmp-lg/9506006
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/151202
dc.subjectComputation and Language
dc.titleAutomatic Extraction of Tagset Mappings from Parallel-Annotated Corpora
dc.typetext

Files

Collections