Skip to main content
    • Aa
    • Aa

A general feature space for automatic verb classification


Lexical semantic classes of verbs play an important role in structuring complex predicate information in a lexicon, thereby avoiding redundancy and enabling generalizations across semantically similar verbs with respect to their usage. Such classes, however, require many person-years of expert effort to create manually, and methods are needed for automatically assigning verbs to appropriate classes. In this work, we develop and evaluate a feature space to support the automatic assignment of verbs into a well-known lexical semantic classification that is frequently used in natural language processing. The feature space is general – applicable to any class distinctions within the target classification; broad – tapping into a variety of semantic features of the classes; and inexpensive – requiring no more than a POS tagger and chunker. We perform experiments using support vector machines (SVMs) with the proposed feature space, demonstrating a reduction in error rate ranging from 48% to 88% over a chance baseline accuracy, across classification tasks of varying difficulty. In particular, we attain performance comparable to or better than that of feature sets manually selected for the particular tasks. Our results show that the approach is generally applicable, and reduces the need for resource-intensive linguistic analysis for each new classification task. We also perform a wide range of experiments to determine the most informative features in the feature space, finding that simple, easily extractable features suffice for good verb classification performance.

Corresponding author
Current affiliation: Interactive Language Technologies Group, Institute for Information Technology, National Research Council Canada, A1330-101 St-Jean-Bosco Street, Gatineau, Quebec, CanadaJ8Y 3G5.
Hide All
S. Abney (1991) Parsing by chunks. In: R. Berwick , S. Abney and C. Tenny (eds.), Principle-Based Parsing. Kluwer Academic.

D. R. Dowty (1991) Thematic proto-roles and argument selection. Language, 67 (3): 547619.

D. Gildea and D. Jurafsky (2002) Automatic labeling of semantic roles. Computational Linguistics, 28 (3): 245288.

N. Habash , B. J. Dorr and D. Traum (2003) Hybrid natural language generation from lexical conceptual structures. Machine Translation, 18 (2): 81128.

M. Lapata and C. Brew (2004) Verb class disambiguation using informative priors. Computational Linguistics, 30 (1): 4573.

P. Merlo and S. Stevenson (2001) Automatic verb classification based on statistical distributions of argument structure. Computational Linguistics, 27 (3): 373408.

A. Oishi and Y. Matsumoto (1997) Detecting the organization of semantic subclasses of Japanese verbs. Int. J. Corpus Linguistics, 2 (1): 6589.

M. Palmer , D. Gildea and P. Kingsbury (2005) The Proposition Bank: An annotated corpus of semantic roles. Computational Linguistics, 31 (1): 71106.

P. Resnik (1996) Selectional constraints: an information-theoretic model and its computational realization. Cognition, 61 (1–2): 127159.

A. Villavicencio (2005) The availability of verb-particle constructions in lexical resources: How much is enough? Computer Speech and Language, Special Issue on Multiword Expressions, 19 (4): 415432.

Recommend this journal

Email your librarian or administrator to recommend adding this journal to your organisation's collection.

Natural Language Engineering
  • ISSN: 1351-3249
  • EISSN: 1469-8110
  • URL: /core/journals/natural-language-engineering
Please enter your name
Please enter a valid email address
Who would you like to send this to? *


Full text views

Total number of HTML views: 2
Total number of PDF views: 12 *
Loading metrics...

Abstract views

Total abstract views: 128 *
Loading metrics...

* Views captured on Cambridge Core between September 2016 - 17th October 2017. This data will be updated every 24 hours.