Automatic Extraction of Subcategorization Frames for Czech

dc.creatorSarkar, Anoop
dc.creatorZeman, Daniel
dc.date2000-09-08
dc.date.accessioned2026-07-07T03:16:32Z
dc.date.available2026-07-07T03:16:32Z
dc.descriptionWe present some novel machine learning techniques for the identification of subcategorization information for verbs in Czech. We compare three different statistical techniques applied to this problem. We show how the learning algorithm can be used to discover previously unknown subcategorization frames from the Czech Prague Dependency Treebank. The algorithm can then be used to label dependents of a verb in the Czech treebank as either arguments or adjuncts. Using our techniques, we ar able to achieve 88% precision on unseen parsed text.
dc.description7 pages. Another version under the name "Learning Verb Subcategorization from Corpora: Counting Frame Subsets", authors: Zeman, Sarkar, in proceedings of LREC 2000, Athens, Greece
dc.identifierhttps://arxiv.org/abs/cs/0009003
dc.identifierhttp://arxiv.org/abs/cs/0009003
dc.identifierProceedings of the 18th International Conference on Computational Linguistics (Coling 2000), Universit
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/30386
dc.subjectComputation and Language
dc.subjectI.2.7, G.3
dc.titleAutomatic Extraction of Subcategorization Frames for Czech
dc.typetext

Files

Collections