Xiao Tang, Tianyu Wang, Chi-Wing Fu
3D hand-mesh reconstruction from RGB images facilitates many applications, including augmented reality (AR). However, this requires not only real-time speed and accurate hand pose and shape but also plausible mesh-image alignment. While existing works already achieve promising results, meeting all three requirements is very challenging. This paper presents a novel pipeline by decoupling the hand-mesh reconstruction task into three stages: a joint stage to predict hand joints and segmentation; a mesh stage to predict a rough hand mesh; and a refine stage to fine-tune it with an offset mesh for mesh-image alignment. With careful design in the network structure and in the loss functions, we can promote high-quality finger-level mesh-image alignment and drive the models together to deliver real-time predictions. Extensive quantitative and qualitative results on benchmark datasets demonstrate that the quality of our results outperforms the state-of-the-art methods on hand-mesh/pose precision and hand-image alignment. In the end, we also showcase several real-time AR scenarios.
| Task | Dataset | Metric | Value | Model |
|---|---|---|---|---|
| Hand | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| Hand | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| Hand | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| Hand | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |
| Pose Estimation | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| Pose Estimation | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| Pose Estimation | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| Pose Estimation | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |
| Hand Pose Estimation | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| Hand Pose Estimation | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| Hand Pose Estimation | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| Hand Pose Estimation | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |
| 3D | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| 3D | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| 3D | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| 3D | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |
| 3D Hand Pose Estimation | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| 3D Hand Pose Estimation | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| 3D Hand Pose Estimation | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| 3D Hand Pose Estimation | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |
| 1 Image, 2*2 Stitchi | FreiHAND | PA-F@15mm | 0.981 | Tang et al. |
| 1 Image, 2*2 Stitchi | FreiHAND | PA-F@5mm | 0.724 | Tang et al. |
| 1 Image, 2*2 Stitchi | FreiHAND | PA-MPJPE | 6.7 | Tang et al. |
| 1 Image, 2*2 Stitchi | FreiHAND | PA-MPVPE | 6.7 | Tang et al. |