arXiv · cs/0009003
Automatic Extraction of Subcategorization Frames for Czech
Abstract
We present some novel machine learning techniques for the identification of subcategorization information for verbs in Czech. We compare three different statistical techniques applied to this problem. We show how the learning algorithm can be used to discover previously unknown subcategorization frames from the Czech Prague Dependency Treebank. The algorithm can then be used to label dependents of a verb in the Czech treebank as either arguments or adjuncts. Using our techniques, we ar able to achieve 88% precision on unseen parsed text.
Explore related subjects
Keep this discovery
Anoop Sarkar, Daniel Zeman. 2000-09-08. Automatic Extraction of Subcategorization Frames for Czech. https://arxiv.org/abs/cs/0009003
Cite the original work for its findings. Save a collection to share your selection of sources.