Abstract
Traditionally, word sense disambiguation (WSD) involves a different context model for each individual word. This paper presents a new approach to WSD using weakly supervised learning. Statistical models are not trained for the contexts of each individual word, but for the similarities between context pairs at category level. The insight is that the correlation regularity between the sense distinction and the context distinction can be captured at category level, independent of individual words. This approach only requires a limited amount of existing annotated training corpus in order to disambiguate the entire vocabulary. A context clustering scheme is developed within the Bayesian framework. A maximum entropy model is then trained to represent the generative probability distribution of context similarities based on heterogeneous features, including trigger words and parsing structures. Statistical annealing is applied to derive the final context clusters by globally fitting the pairwise context similarity distribution. Benchmarking shows that this new approach significantly outperforms the existing WSD systems in the unsupervised category, and rivals supervised WSD systems.
| Original language | English |
|---|---|
| Pages | 187-190 |
| Number of pages | 4 |
| State | Published - 2004 |
| Event | 3rd International Workshop on the Evaluation of Systems for the Semantic Analysis of Text, SENSEVAL@ACL 2004 - Barcelona, Spain Duration: Jul 25 2004 → Jul 26 2004 |
Conference
| Conference | 3rd International Workshop on the Evaluation of Systems for the Semantic Analysis of Text, SENSEVAL@ACL 2004 |
|---|---|
| Country/Territory | Spain |
| City | Barcelona |
| Period | 07/25/04 → 07/26/04 |
Fingerprint
Dive into the research topics of 'Context clustering for word sense disambiguation based on modeling pairwise context similarities'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver