Skip to Main Content (Press Enter)

Logo CNR
  • ×
  • Home
  • People
  • Outputs
  • Organizations
  • Expertise & Skills

UNI-FIND
Logo CNR

|

UNI-FIND

cnr.it
  • ×
  • Home
  • People
  • Outputs
  • Organizations
  • Expertise & Skills
  1. Outputs

Automatic web page categorization by link and context analysis

Conference Paper
Publication Date:
1999
abstract:
Assistance in retrieving documents on the World Wide Web is provided either by search engines, through keyword-based queries, or by catalogues, which organize documents into hierarchical collections. Maintaining catalogues manually is becoming increasingly difficult, due to the sheer amount of material on the Web; it is thus becoming necessary to resort to techniques for the automatic classification of documents. Automatic classification is traditionally performed by extracting the information for representing a document ("indexing") from the document itself. The paper describes the novel technique of categorization by context, which instead extracts useful information for classifying a document from the context where a URL referring to it appears. We present the results of experimenting with Theseus, a classifier that exploits this technique.
Iris type:
04.01 Contributo in Atti di convegno
Keywords:
Information search and retrieval
List of contributors:
Sebastiani, Fabrizio
Authors of the University:
SEBASTIANI FABRIZIO
Handle:
https://iris.cnr.it/handle/20.500.14243/387433
Full Text:
https://iris.cnr.it//retrieve/handle/20.500.14243/387433/69529/prod_407747-doc_142945.pdf
  • Use of cookies

Powered by VIVO | Designed by Cineca | 26.5.0.0 | Sorgente dati: PREPROD (Ribaltamento disabilitato)