Improving term extraction by system combination using boosting

Jordi Vivaldi, Lluis Marques, Horacio Rodríguez

Research output: Chapter in Book/Report/Conference proceedingConference contribution

31 Citations (Scopus)

Abstract

Term extraction is the task of automatically detecting, from textual corpora, lexical units that designate concepts in thematically restricted domains (e.g. medicine). Current systems for term extraction integrate linguistic and statistical cues to perform the detection of terms. The best results have been obtained when some kind of combination of simple base term extractors is performed [14]. In this paper it is shown that this combination can be further improved by posing an additional learning problem of how to find the best combination of base term extractors. Empirical results, using AdaBoost in the metalearning step, show that the ensemble constructed surpasses the performance of all individual extractors and simple voting schemes, obtaining significantly better accuracy figures at all levels of recall.

Original languageEnglish
Title of host publicationLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
PublisherSpringer Verlag
Pages515-526
Number of pages12
Volume2167
ISBN (Print)3540425365, 9783540425366
Publication statusPublished - 2001
Externally publishedYes
Event12th European Conference on Machine Learning, ECML 2001 - Freiburg, Germany
Duration: 5 Sep 20017 Sep 2001

Publication series

NameLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volume2167
ISSN (Print)03029743
ISSN (Electronic)16113349

Other

Other12th European Conference on Machine Learning, ECML 2001
CountryGermany
CityFreiburg
Period5/9/017/9/01

    Fingerprint

ASJC Scopus subject areas

  • Computer Science(all)
  • Theoretical Computer Science

Cite this

Vivaldi, J., Marques, L., & Rodríguez, H. (2001). Improving term extraction by system combination using boosting. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 2167, pp. 515-526). (Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics); Vol. 2167). Springer Verlag.