In this paper, we investigate the usefulness of the reject option in text categorisation systems. The reject option is introduced by allowing a text classifier to withhold the decision of assigning or not a document to any subset of categories, for which the decision is considered not sufficiently reliable. To automatically handle rejections, a two-stage classifier architecture is used, in which documents rejected at the first stage are automatically classified at the second stage, so that no rejections eventually remain. The performance improvement achievable by using the reject option is assessed on a real text categorisation task, using the well known Reuters data set.
Scheda prodotto non validato
Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo
Titolo: | A Two-Stage Classifier with Reject Option for Text Categorisation |
Autori: | |
Data di pubblicazione: | 2004 |
Handle: | http://hdl.handle.net/11584/15625 |
ISBN: | 978-3-540-22570-6 |
Tipologia: | 4.1 Contributo in Atti di convegno |