Manual Annotation of Translational Equivalence: The Blinker Project
| dc.creator | Melamed, I. Dan | |
| dc.date | 1998-05-08 | |
| dc.date | 1998-05-11 | |
| dc.date.accessioned | 2026-07-07T02:36:14Z | |
| dc.date.available | 2026-07-07T02:36:14Z | |
| dc.description | Bilingual annotators were paid to link roughly sixteen thousand corresponding words between on-line versions of the Bible in modern French and modern English. These annotations are freely available to the research community from http://www.cis.upenn.edu/~melamed . The annotations can be used for several purposes. First, they can be used as a standard data set for developing and testing translation lexicons and statistical translation models. Second, researchers in lexical semantics will be able to mine the annotations for insights about cross-linguistic lexicalization patterns. Third, the annotations can be used in research into certain recently proposed methods for monolingual word-sense disambiguation. This paper describes the annotated texts, the specially-designed annotation tool, and the strategies employed to increase the consistency of the annotations. The annotation process was repeated five times by different annotators. Inter-annotator agreement rates indicate that the annotations are reasonably reliable and that the method is easy to replicate. | |
| dc.identifier | https://arxiv.org/abs/cmp-lg/9805005 | |
| dc.identifier | http://arxiv.org/abs/cmp-lg/9805005 | |
| dc.identifier.uri | http://salesiana.dossiersoluciones.com/handle/123456789/15839 | |
| dc.subject | Computation and Language | |
| dc.title | Manual Annotation of Translational Equivalence: The Blinker Project | |
| dc.type | text |