Abstract
Previous studies have shown that one-class SVM is a rather weak learning method for text categorization problems. This paper points out that the poor performance observed before is largely due to the fact that the standard term weighting schemes are inadequate for one-class SVMs. We propose several representation modifications, and demonstrate empirically that, with the proposed document representation, the performance of one-class SVM, although trained on only small portion of positive examples, can reach up to 95% of that of two-class SVM trained on the whole labeled dataset.
| Original language | English |
|---|---|
| Pages (from-to) | 489-500 |
| Number of pages | 12 |
| Journal | Lecture Notes in Computer Science |
| Volume | 3201 |
| DOIs | |
| State | Published - 2004 |
| Event | 15th European Conference on Machine Learning, ECML 2004 - Pisa, Italy Duration: Sep 20 2004 → Sep 24 2004 |
Fingerprint
Dive into the research topics of 'Document representation for one-class SVM'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver