Skip to main navigation Skip to search Skip to main content

Temporally enhanced image object proposals for videos

  • Nanyang Technological University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Despite the recent success of image object proposals (IOPs) for image applications, the per-frame IOPs are also important for video applications. However, the existing IOPs are extracted from each frame separately and may exhibit inconsistencies across the frames. In this paper, we propose to improve the existing IOPs by enforcing the temporal consistency through a video sequence in an on-line manner. To achieve this, we propose a novel spatio-temporal objectness measure considering both the frame level objectness as well as the temporal consistency across frames. An on-line dynamic programing technique is proposed to efficiently compute such spatio-temporal objectness. In addition, compared with the spatio-temporal video object proposals(VOPs), the proposed method supports on-line applications and provides more accurate per-frame localizations. Experiments on benchmark datasets validate its superior performance compared with the existing IOPs and VOPs.

Original languageEnglish
Title of host publication2017 IEEE International Conference on Multimedia and Expo, ICME 2017
PublisherIEEE Computer Society
Pages445-450
Number of pages6
ISBN (Electronic)9781509060672
DOIs
StatePublished - Aug 28 2017
Event2017 IEEE International Conference on Multimedia and Expo, ICME 2017 - Hong Kong, Hong Kong
Duration: Jul 10 2017Jul 14 2017

Publication series

NameProceedings - IEEE International Conference on Multimedia and Expo
ISSN (Print)1945-7871
ISSN (Electronic)1945-788X

Conference

Conference2017 IEEE International Conference on Multimedia and Expo, ICME 2017
Country/TerritoryHong Kong
CityHong Kong
Period07/10/1707/14/17

Keywords

  • Object proposals
  • On-line
  • Video

Fingerprint

Dive into the research topics of 'Temporally enhanced image object proposals for videos'. Together they form a unique fingerprint.

Cite this