Skip to Main Content (Press Enter)

Logo CNR
  • ×
  • Home
  • Persone
  • Pubblicazioni
  • Strutture
  • Competenze

UNI-FIND
Logo CNR

|

UNI-FIND

cnr.it
  • ×
  • Home
  • Persone
  • Pubblicazioni
  • Strutture
  • Competenze
  1. Pubblicazioni

Automatic expansion of domain-specific lexicons by term categorization

Articolo
Data di Pubblicazione:
2006
Abstract:
We discuss an approach to the automatic expansion of domain-specific lexicons, i.e., to the problem of extending, for each ci in a predefined set C = {c1, . . . , cm} of semantic domains, an initial lexicon Li 0 into a larger lexicon Li 1. Our approach relies on term categorization, defined as the task of labeling previously unlabeled terms according to a predefined set of domains. We approach this as a supervised learning problem, in which term classifiers are built using the initial lexicons as training data. Dually to classic text categorization tasks, in which documents are represented as vectors in a space of terms, we represent terms as vectors in a space of documents. We present the results of a number of experiments in which we use a boosting-based learning device for training our term classifiers. We test the effectiveness of our method by using WordNetDomains, a well-known large set of domain-specific lexicons, as a benchmark. Our experiments are performed using the documents in the Reuters Corpus Volume 1 as 'implicit' representations for our terms.
Tipologia CRIS:
01.01 Articolo in rivista
Keywords:
I.5.2 Classifier design and evaluation; Lexicons
Elenco autori:
Avancini, HENRI HECTOR; Sebastiani, Fabrizio
Autori di Ateneo:
SEBASTIANI FABRIZIO
Link alla scheda completa:
https://iris.cnr.it/handle/20.500.14243/62927
Pubblicato in:
ACM TRANSACTIONS ON SPEECH AND LANGUAGE PROCESSING
Journal
  • Dati Generali

Dati Generali

URL

http://dl.acm.org/citation.cfm?doid=1138379.1138380
  • Utilizzo dei cookie

Realizzato con VIVO | Designed by Cineca | 26.5.0.0 | Sorgente dati: PREPROD (Ribaltamento disabilitato)