The browser you are using is not supported by this website. All versions of Internet Explorer are no longer supported, either by us or Microsoft (read more here: https://www.microsoft.com/en-us/microsoft-365/windows/end-of-ie-support).

Please use a modern browser to fully experience our website, such as the newest versions of Edge, Chrome, Firefox or Safari etc.

DESIRE Toolkit Components - Matcher

Author

  • Anders Ardö

Summary, in English

The Matcher tool implements a subject classification process using a subject-specific thesaurus by which terms are intellectually mapped to categories or subject classes. The classification process is made up of several steps. First, the document to be classified is fetched. Text is extracted from this document, and all thesaurus terms are matched to it. Some heuristic processing rules are applied to the results from the matching process. Finally, the outcome is formatted either for presentation or for storing in a database.

Publishing year

2000

Language

English

Document type

Report

Publisher

DESIRE EU project

Topic

  • Electrical Engineering, Electronic Engineering, Information Engineering

Keywords

  • focussed web crawling
  • Automated classification

Status

Published