Abstract.
Technical terms (henceforth called terms ), are important elements for digital libraries. In this paper we present a domain-independent method for the automatic extraction of multi-word terms, from machine-readable special language corpora. The method, (C-value/NC-value ), combines linguistic and statistical information. The first part, C-value, enhances the common statistical measure of frequency of occurrence for term extraction, making it sensitive to a particular type of multi-word terms, the nested terms. The second part, NC-value, gives: 1) a method for the extraction of term context words (words that tend to appear with terms); 2) the incorporation of information from term context words to the extraction of terms.
Similar content being viewed by others
Author information
Authors and Affiliations
Additional information
Received: 17 December 1998 / Revised: 19 May 1999
Rights and permissions
About this article
Cite this article
Frantzi, K., Ananiadou, S. & Mima, H. Automatic recognition of multi-word terms:. the C-value/NC-value method . Int J Digit Libr 3, 115–130 (2000). https://doi.org/10.1007/s007999900023
Issue Date:
DOI: https://doi.org/10.1007/s007999900023