TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/SSP-Net: Scalable Sequential Pyramid Networks for Real-Tim...

SSP-Net: Scalable Sequential Pyramid Networks for Real-Time 3D Human Pose Regression

Diogo Luvizon, Hedi Tabia, David Picard

2020-09-043D Human Pose EstimationregressionPose Estimation3D Pose Estimation
PaperPDF

Abstract

In this paper we propose a highly scalable convolutional neural network, end-to-end trainable, for real-time 3D human pose regression from still RGB images. We call this approach the Scalable Sequential Pyramid Networks (SSP-Net) as it is trained with refined supervision at multiple scales in a sequential manner. Our network requires a single training procedure and is capable of producing its best predictions at 120 frames per second (FPS), or acceptable predictions at more than 200 FPS when cut at test time. We show that the proposed regression approach is invariant to the size of feature maps, allowing our method to perform multi-resolution intermediate supervisions and reaching results comparable to the state-of-the-art with very low resolution feature maps. We demonstrate the accuracy and the effectiveness of our method by providing extensive experiments on two of the most important publicly available datasets for 3D pose estimation, Human3.6M and MPI-INF-3DHP. Additionally, we provide relevant insights about our decisions on the network architecture and show its flexibility to meet the best precision-speed compromise.

Results

TaskDatasetMetricValueModel
3D Human Pose EstimationMPI-INF-3DHPAUC44.3SSP-Net
3D Human Pose EstimationMPI-INF-3DHPMPJPE96.8SSP-Net
3D Human Pose EstimationMPI-INF-3DHPPCK83.2SSP-Net
Pose EstimationMPI-INF-3DHPAUC44.3SSP-Net
Pose EstimationMPI-INF-3DHPMPJPE96.8SSP-Net
Pose EstimationMPI-INF-3DHPPCK83.2SSP-Net
3DMPI-INF-3DHPAUC44.3SSP-Net
3DMPI-INF-3DHPMPJPE96.8SSP-Net
3DMPI-INF-3DHPPCK83.2SSP-Net
1 Image, 2*2 StitchiMPI-INF-3DHPAUC44.3SSP-Net
1 Image, 2*2 StitchiMPI-INF-3DHPMPJPE96.8SSP-Net
1 Image, 2*2 StitchiMPI-INF-3DHPPCK83.2SSP-Net

Related Papers

Language Integration in Fine-Tuning Multimodal Large Language Models for Image-Based Regression2025-07-20$π^3$: Scalable Permutation-Equivariant Visual Geometry Learning2025-07-17Revisiting Reliability in the Reasoning-based Pose Estimation Benchmark2025-07-17DINO-VO: A Feature-based Visual Odometry Leveraging a Visual Foundation Model2025-07-17From Neck to Head: Bio-Impedance Sensing for Head Pose Estimation2025-07-17AthleticsPose: Authentic Sports Motion Dataset on Athletic Field and Evaluation of Monocular 3D Pose Estimation Ability2025-07-17Neural Network-Guided Symbolic Regression for Interpretable Descriptor Discovery in Perovskite Catalysts2025-07-16Imbalanced Regression Pipeline Recommendation2025-07-16