TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/xR-EgoPose: Egocentric 3D Human Pose from an HMD Camera

xR-EgoPose: Egocentric 3D Human Pose from an HMD Camera

Denis Tome, Patrick Peluse, Lourdes Agapito, Hernan Badino

2019-07-23ICCV 2019 10Egocentric Pose EstimationPose Estimation
PaperPDF

Abstract

We present a new solution to egocentric 3D body pose estimation from monocular images captured from a downward looking fish-eye camera installed on the rim of a head mounted virtual reality device. This unusual viewpoint, just 2 cm. away from the user's face, leads to images with unique visual appearance, characterized by severe self-occlusions and strong perspective distortions that result in a drastic difference in resolution between lower and upper body. Our contribution is two-fold. Firstly, we propose a new encoder-decoder architecture with a novel dual branch decoder designed specifically to account for the varying uncertainty in the 2D joint locations. Our quantitative evaluation, both on synthetic and real-world datasets, shows that our strategy leads to substantial improvements in accuracy over state of the art egocentric pose estimation approaches. Our second contribution is a new large-scale photorealistic synthetic dataset - xR-EgoPose - offering 383K frames of high quality renderings of people with a diversity of skin tones, body shapes, clothing, in a variety of backgrounds and lighting conditions, performing a range of actions. Our experiments show that the high variability in our new synthetic training corpus leads to good generalization to real world footage and to state of the art results on real world datasets with ground truth. Moreover, an evaluation on the Human3.6M benchmark shows that the performance of our method is on par with top performing approaches on the more classic problem of 3D human pose from a third person viewpoint.

Results

TaskDatasetMetricValueModel
3D Human Pose EstimationGlobalEgoMocap Test DatasetAverage MPJPE (mm)112xR-egopose
3D Human Pose EstimationGlobalEgoMocap Test DatasetPA-MPJPE87.2xR-egopose
3D Human Pose EstimationSceneEgoAverage MPJPE (mm)241.3xR-egopose
3D Human Pose EstimationSceneEgoPA-MPJPE133.9xR-egopose
Pose EstimationGlobalEgoMocap Test DatasetAverage MPJPE (mm)112xR-egopose
Pose EstimationGlobalEgoMocap Test DatasetPA-MPJPE87.2xR-egopose
Pose EstimationSceneEgoAverage MPJPE (mm)241.3xR-egopose
Pose EstimationSceneEgoPA-MPJPE133.9xR-egopose
3DGlobalEgoMocap Test DatasetAverage MPJPE (mm)112xR-egopose
3DGlobalEgoMocap Test DatasetPA-MPJPE87.2xR-egopose
3DSceneEgoAverage MPJPE (mm)241.3xR-egopose
3DSceneEgoPA-MPJPE133.9xR-egopose
1 Image, 2*2 StitchiGlobalEgoMocap Test DatasetAverage MPJPE (mm)112xR-egopose
1 Image, 2*2 StitchiGlobalEgoMocap Test DatasetPA-MPJPE87.2xR-egopose
1 Image, 2*2 StitchiSceneEgoAverage MPJPE (mm)241.3xR-egopose
1 Image, 2*2 StitchiSceneEgoPA-MPJPE133.9xR-egopose

Related Papers

$π^3$: Scalable Permutation-Equivariant Visual Geometry Learning2025-07-17Revisiting Reliability in the Reasoning-based Pose Estimation Benchmark2025-07-17DINO-VO: A Feature-based Visual Odometry Leveraging a Visual Foundation Model2025-07-17From Neck to Head: Bio-Impedance Sensing for Head Pose Estimation2025-07-17AthleticsPose: Authentic Sports Motion Dataset on Athletic Field and Evaluation of Monocular 3D Pose Estimation Ability2025-07-17SpatialTrackerV2: 3D Point Tracking Made Easy2025-07-16SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation2025-07-16Efficient Calisthenics Skills Classification through Foreground Instance Selection and Depth Estimation2025-07-16