• Home
  • Search
  • Markerless multi-view 3D human pose estimation: A survey
  • Open Access IconOpen Access
  • Cite Icon26
  • https://doi.org/10.1016/j.imavis.2025.105437Copy DOI Icon

Markerless multi-view 3D human pose estimation: A survey

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

3D human pose estimation aims to reconstruct the human skeleton of all the individuals in a scene by detecting several body joints. The creation of accurate and efficient methods is required for several real-world applications including animation, human–robot interaction, surveillance systems or sports, among many others. However, several obstacles such as occlusions, random camera perspectives, or the scarcity of 3D labelled data, have been hampering the models’ performance and limiting their deployment in real-world scenarios. The higher availability of cameras has led researchers to explore multi-view solutions due to the advantage of being able to exploit different perspectives to reconstruct the pose. Most existing reviews focus mainly on monocular 3D human pose estimation and a comprehensive survey only on multi-view approaches to determine the 3D pose has been missing since 2012. Thus, the goal of this survey is to fill that gap and present an overview of the methodologies related to 3D pose estimation in multi-view settings, understand what were the strategies found to address the various challenges and also, identify their limitations. According to the reviewed articles, it was possible to find that most methods are fully-supervised approaches based on geometric constraints. Nonetheless, most of the methods suffer from 2D pose mismatches, to which the incorporation of temporal consistency and depth information have been suggested to reduce the impact of this limitation, besides working directly with 3D features can completely surpass this problem but at the expense of higher computational complexity. Models with lower supervision levels were identified to overcome some of the issues related to 3D pose, particularly the scarcity of labelled datasets. Therefore, no method is yet capable of solving all the challenges associated with the reconstruction of the 3D pose. Due to the existing trade-off between complexity and performance, the best method depends on the application scenario. Therefore, further research is still required to develop an approach capable of quickly inferring a highly accurate 3D pose with bearable computation cost. To this goal, techniques such as active learning, methods that learn with a low level of supervision, the incorporation of temporal consistency, view selection, estimation of depth information and multi-modal approaches might be interesting strategies to keep in mind when developing a new methodology to solve this task. • First review only on multi-view, multi-modal methods to estimate 3D pose since 2012. • Multi-view allows capturing the full body geometry, making 3D pose estimation easier. • Real-world applications include sports, broadcasting, rehabilitation or animation. • Finding a fast, accurate method with low computational cost remains a challenge. • Multi-modal methods or view selection can lead to an efficient and effective model.

Similar Papers
  • Research Article
  • Citations31

Weakly-supervised pre-training for 3D human pose estimation via perspective knowledge

  • Mar 05, 2023
  • Pattern Recognition
  • Zhongwei Qiu +3
  • Conference Article
  • Citations4

Sparse representation and convolutional neural networks for 3D human pose estimation

  • Dec 01, 2017
  • Hassan Alikarami +2
  • Research Article
  • Citations52

Boosting Monocular 3D Human Pose Estimation With Part Aware Attention.

  • Jan 01, 2022
  • IEEE Transactions on Image Processing
  • Youze Xue +4
  • Research Article
  • Citations381

Deep 3D human pose estimation: A review

  • May 24, 2021
  • Computer Vision and Image Understanding
  • Jinbao Wang +6
  • Conference Article
  • Citations135

Iterated Second-Order Label Sensitive Pooling for 3D Human Pose Estimation

  • Jun 01, 2014
  • Catalin Ionescu +2
  • Conference Article
  • Citations2

3D Body Pose and Shape Estimation from Multi-View Images With Limb Geometric Constraint

  • Sep 01, 2019
  • Zixuan Gai +2
  • Conference Article
  • Citations3

Viewpoint Selection for DermDrone using Deep Reinforcement Learning

  • Oct 12, 2021
  • Mojtaba Ahangar Arzati +1
  • Research Article

Cross-Constrained Progressive Inference for 3D Hand Pose Estimation with Dynamic Observer-Decision-Adjuster Networks

  • Mar 24, 2024
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Zhehan Kan +4
  • Conference Article
  • Citations42

ActiveMoCap: Optimized Viewpoint Selection for Active Human Motion Capture

  • Jun 01, 2020
  • Sena Kiciroglu +4
  • Video Transcripts

On the Robustness of 3D Human Pose Estimation

  • Dec 29, 2020
  • Underline Science Inc.
  • Yan Huang +2
  • Book Chapter
  • Citations1

Towards Generalization of 3D Human Pose Estimation in the Wild

  • Jan 01, 2021
  • Renato Baptista +3
  • Research Article
  • Citations17

Enhancement of human 3D pose estimation using a novel concept of depth prediction with pose alignment from a single 2D image

  • Jul 26, 2022
  • Computers & Graphics
  • Mohit Kushwaha +2
  • Book Chapter
  • Citations14

Unsupervised Cross-Modal Alignment for Multi-person 3D Pose Estimation

  • Jan 01, 2020
  • Jogendra Nath Kundu +4
  • Research Article
  • Citations46

Geometry-Driven Self-Supervised Method for 3D Human Pose Estimation

  • Apr 03, 2020
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Yang Li +5
  • Research Article
  • Citations1

Local-Global Feature Fusion for Enhancing 3D Human Pose Estimation

  • Jan 01, 2025
  • IEEE Transactions on Circuits and Systems for Video Technology
  • Yuanhong Zhong +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.