Skip to main navigation Skip to search Skip to main content

Fusion of 3D-LIDAR and camera data for scene parsing

  • Nanyang Technological University
  • DSO National Laboratory, Singapore

Research output: Contribution to journalArticlepeer-review

55 Scopus citations

Abstract

Fusion of information gathered from multiple sources is essential to build a comprehensive situation picture for autonomous ground vehicles. In this paper, an approach which performs scene parsing and data fusion for a 3D-LIDAR scanner (Velodyne HDL-64E) and a video camera is described. First of all, a geometry segmentation algorithm is proposed for detection of obstacles and ground areas from data collected by the Velodyne scanner. Then, corresponding image collected by the video camera is classified patch by patch into more detailed categories. After that, parsing result of each frame is obtained by fusing result of Velodyne data and that of image using the fuzzy logic inference framework. Finally, parsing results of consecutive frames are smoothed by the Markov random field based temporal fusion method. The proposed approach has been evaluated with datasets collected by our autonomous ground vehicle testbed in both rural and urban areas. The fused results are more reliable than that acquired via analysis of only images or Velodyne data.

Original languageEnglish
Pages (from-to)165-183
Number of pages19
JournalJournal of Visual Communication and Image Representation
Volume25
Issue number1
DOIs
StatePublished - Jan 2014

Keywords

  • Camera
  • Fuzzy logic
  • MRF
  • Object detection
  • RGBD
  • Scene parsing
  • Temporal fusion
  • Velodyne scanner

Fingerprint

Dive into the research topics of 'Fusion of 3D-LIDAR and camera data for scene parsing'. Together they form a unique fingerprint.

Cite this