Skip to main navigation Skip to search Skip to main content

Exploiting spatial-temporal relationships for 3D pose estimation via graph convolutional networks

  • Yujun Cai
  • , Liuhao Ge
  • , Jun Liu
  • , Jianfei Cai
  • , Tat Jen Cham
  • , Junsong Yuan
  • , Nadia Magnenat Thalmann
  • Nanyang Technological University
  • Monash University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

559 Scopus citations

Abstract

Despite great progress in 3D pose estimation from single-view images or videos, it remains a challenging task due to the substantial depth ambiguity and severe self-occlusions. Motivated by the effectiveness of incorporating spatial dependencies and temporal consistencies to alleviate these issues, we propose a novel graph-based method to tackle the problem of 3D human body and 3D hand pose estimation from a short sequence of 2D joint detections. Particularly, domain knowledge about the human hand (body) configurations is explicitly incorporated into the graph convolutional operations to meet the specific demand of the 3D pose estimation. Furthermore, we introduce a local-to-global network architecture, which is capable of learning multi-scale features for the graph-based representations. We evaluate the proposed method on challenging benchmark datasets for both 3D hand pose estimation and 3D body pose estimation. Experimental results show that our method achieves state-of-the-art performance on both tasks.

Original languageEnglish
Title of host publicationProceedings - 2019 International Conference on Computer Vision, ICCV 2019
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages2272-2281
Number of pages10
ISBN (Electronic)9781728148038
DOIs
StatePublished - Oct 2019
Event17th IEEE/CVF International Conference on Computer Vision, ICCV 2019 - Seoul, Korea, Republic of
Duration: Oct 27 2019Nov 2 2019

Publication series

NameProceedings of the IEEE International Conference on Computer Vision
ISSN (Print)1550-5499

Conference

Conference17th IEEE/CVF International Conference on Computer Vision, ICCV 2019
Country/TerritoryKorea, Republic of
CitySeoul
Period10/27/1911/2/19

Fingerprint

Dive into the research topics of 'Exploiting spatial-temporal relationships for 3D pose estimation via graph convolutional networks'. Together they form a unique fingerprint.

Cite this