Using the DIFF Command for Natural Language Processing

dc.creatorMurata, Masaki
dc.creatorIsahara, Hitoshi
dc.date2002-08-13
dc.date.accessioned2026-07-07T03:18:47Z
dc.date.available2026-07-07T03:18:47Z
dc.descriptionDiff is a software program that detects differences between two data sets and is useful in natural language processing. This paper shows several examples of the application of diff. They include the detection of differences between two different datasets, extraction of rewriting rules, merging of two different datasets, and the optimal matching of two different data sets. Since diff comes with any standard UNIX system, it is readily available and very easy to use. Our studies showed that diff is a practical tool for research into natural language processing.
dc.description10 pages. Computation and Language. This paper is the rough English translation of our Japanese papar
dc.identifierhttps://arxiv.org/abs/cs/0208020
dc.identifierhttp://arxiv.org/abs/cs/0208020
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/31258
dc.subjectComputation and Language
dc.subjectH.3.3; I.2.7
dc.titleUsing the DIFF Command for Natural Language Processing
dc.typetext

Files

Collections