☆ 4.7 Article Proceedings Paper

Robust 3D Hand Pose Estimation From Single Depth Images Using Multi-View CNNs

IEEE TRANSACTIONS ON IMAGE PROCESSING (2018)

Journal

IEEE TRANSACTIONS ON IMAGE PROCESSING

Volume 27, Issue 9, Pages 4422-4436

Publisher

IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC

DOI: 10.1109/TIP.2018.2834824

Keywords

3D hand pose estimation; convolutional neural networks; multi-view CNNs

Funding

BeingTogether Centre
National Research Foundation, Prime Minister's Office, Singapore under its International Research Centres in Singapore Funding Initiative
Singapore Ministry of Education Academic Research Fund [MOE2015-T2-2-114]
Microsoft Research Asia
University at Buffalo

Ask authors/readers for more resources

Protocol

Community support

Reagent

Community support

Abstract

Articulated hand pose estimation is one of core technologies in human-computer interaction. Despite the recent progress, most existing methods still cannot achieve satisfactory performance, partly due to the difficulty of the embedded high-dimensional nonlinear regression problem. Most existing data-driven methods directly regress 3D hand pose from 2D depth image, which cannot fully utilize the depth information. In this paper, we propose a novel multi-view convolutional neural network (CNN)-based approach for 3D hand pose estimation. To better exploit 3D information in the depth image, we project the point cloud generated from the query depth image onto multiple views of two projection settings and integrate them for more robust estimation. Multi-view CNNs are trained to learn the mapping from projected images to heat-maps, which reflect probability distributions of joints on each view. These multi-view heat-maps are then fused to estimate the optimal 3D hand pose with learned pose priors, and the unreliable information in multi-view heat-maps is suppressed using a view selection method. Experimental results show that the proposed method is superior to the state-of-the-art methods on two challenging data sets. Furthermore, a cross-data set experiment also validates that our proposed approach has good generalization ability.

Authors

I am an author on this paper

Click your name to claim this paper and add it to your profile.

Reviews

Primary Rating

4.7

Not enough ratings

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recurrent 3D Hand Pose Estimation Using Cascaded Pose-Guided 3D Alignments

Xiaoming Deng, Dexin Zuo, Yinda Zhang, Zhaopeng Cui, Jian Cheng, Ping Tan, Liang Chang, Marc Pollefeys, Sean Fanello, Hongan Wang

Summary: This paper investigates the impact of view-independent features on 3D hand pose estimation from a single depth image, and proposes a novel recurrent neural network model for 3D hand pose estimation. The model uses a cascaded 3D pose-guided alignment strategy for view-independent feature extraction and a recurrent hand pose module for modeling the dependencies among sequential aligned features. Experiments show that this method significantly improves the state-of-the-art accuracy on popular benchmarks with simple yet efficient alignment and network architectures.

IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE (2023)