Abstract
Finding useful information from large multimodal document collections such as the WWW without encountering numerous false positives poses a challenge to multimedia information retrieval systems (MMIR). This research addresses the problem of finding pictures. The fact that images do not appear in isolation, but rather with accompanying, collateral text is exploited. Taken independently, existing techniques for picture retrieval using (i) text-based and (ii) image-based methods have several limitations. This research presents a general model for multimodal information retrieval that addresses the following issues: (i) users' information need, (ii) expressing information need through composite, multimodal queries, and (iii) determining the most appropriate weighted combination of indexing techniques in order to best satisfy information need. A machine learning approach is proposed for the latter. The focus is on improving precision and recall in a MMIR system by optimally combining text and image similarity. Experiments are presented which demonstrate the utility of individual indexing systems in improving overall average precision.
| Original language | English |
|---|---|
| Pages (from-to) | 245-275 |
| Number of pages | 31 |
| Journal | Information Retrieval Journal |
| Volume | 2 |
| Issue number | 2-3 |
| DOIs | |
| State | Published - 2000 |
Keywords
- Content-based retrieval
- Image indexing
- Multimedia information retrieval
- Multimodal query processing
- Text indexing
Fingerprint
Dive into the research topics of 'Intelligent indexing and semantic retrieval of multimodal documents'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver