I^2R-Net: Intra- and Inter-Human Relation Network for Multi-Person Pose Estimation

Yiwei Ding, Wenjin Deng, Yinglin Zheng, PengFei Liu, Meihong Wang, Xuan Cheng, Jianmin Bao, Dong Chen, Ming Zeng

2022-06-22Pose Estimation Multi-Person Pose Estimation

Abstract

In this paper, we present the Intra- and Inter-Human Relation Networks (I^2R-Net) for Multi-Person Pose Estimation. It involves two basic modules. First, the Intra-Human Relation Module operates on a single person and aims to capture Intra-Human dependencies. Second, the Inter-Human Relation Module considers the relation between multiple instances and focuses on capturing Inter-Human interactions. The Inter-Human Relation Module can be designed very lightweight by reducing the resolution of feature map, yet learn useful relation information to significantly boost the performance of the Intra-Human Relation Module. Even without bells and whistles, our method can compete or outperform current competition winners. We conduct extensive experiments on COCO, CrowdPose, and OCHuman datasets. The results demonstrate that the proposed model surpasses all the state-of-the-art methods. Concretely, the proposed method achieves 77.4% AP on CrowPose dataset and 67.8% AP on OCHuman dataset respectively, outperforming existing methods by a large margin. Additionally, the ablation study and visualization analysis also prove the effectiveness of our model.

Results

Task	Dataset	Metric	Value	Model
Pose Estimation	COCO (Common Objects in Context)	AP	77.3	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	COCO (Common Objects in Context)	AP50	91	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	COCO (Common Objects in Context)	AP75	83.6	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	COCO (Common Objects in Context)	APL	84.5	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	COCO (Common Objects in Context)	APM	73	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	COCO (Common Objects in Context)	AR	82.1	I²R-Net (1st stage:HRFormer-B)
Pose Estimation	CrowdPose	AP Easy	83.8	I²R-Net (1st stage: HRFormer-B)
Pose Estimation	CrowdPose	AP Hard	69.3	I²R-Net (1st stage: HRFormer-B)
Pose Estimation	CrowdPose	AP Medium	78.1	I²R-Net (1st stage: HRFormer-B)
Pose Estimation	CrowdPose	mAP @0.5:0.95	77.4	I²R-Net (1st stage: HRFormer-B)
Pose Estimation	OCHuman	AP50	85	I²R-Net (1st stage:TransPose-H)
Pose Estimation	OCHuman	AP75	72.8	I²R-Net (1st stage:TransPose-H)
Pose Estimation	OCHuman	Validation AP	67.8	I²R-Net (1st stage:TransPose-H)
3D	COCO (Common Objects in Context)	AP	77.3	I²R-Net (1st stage:HRFormer-B)
3D	COCO (Common Objects in Context)	AP50	91	I²R-Net (1st stage:HRFormer-B)
3D	COCO (Common Objects in Context)	AP75	83.6	I²R-Net (1st stage:HRFormer-B)
3D	COCO (Common Objects in Context)	APL	84.5	I²R-Net (1st stage:HRFormer-B)
3D	COCO (Common Objects in Context)	APM	73	I²R-Net (1st stage:HRFormer-B)
3D	COCO (Common Objects in Context)	AR	82.1	I²R-Net (1st stage:HRFormer-B)
3D	CrowdPose	AP Easy	83.8	I²R-Net (1st stage: HRFormer-B)
3D	CrowdPose	AP Hard	69.3	I²R-Net (1st stage: HRFormer-B)
3D	CrowdPose	AP Medium	78.1	I²R-Net (1st stage: HRFormer-B)
3D	CrowdPose	mAP @0.5:0.95	77.4	I²R-Net (1st stage: HRFormer-B)
3D	OCHuman	AP50	85	I²R-Net (1st stage:TransPose-H)
3D	OCHuman	AP75	72.8	I²R-Net (1st stage:TransPose-H)
3D	OCHuman	Validation AP	67.8	I²R-Net (1st stage:TransPose-H)
Multi-Person Pose Estimation	CrowdPose	AP Easy	83.8	I²R-Net (1st stage: HRFormer-B)
Multi-Person Pose Estimation	CrowdPose	AP Hard	69.3	I²R-Net (1st stage: HRFormer-B)
Multi-Person Pose Estimation	CrowdPose	AP Medium	78.1	I²R-Net (1st stage: HRFormer-B)
Multi-Person Pose Estimation	CrowdPose	mAP @0.5:0.95	77.4	I²R-Net (1st stage: HRFormer-B)
Multi-Person Pose Estimation	OCHuman	AP50	85	I²R-Net (1st stage:TransPose-H)
Multi-Person Pose Estimation	OCHuman	AP75	72.8	I²R-Net (1st stage:TransPose-H)
Multi-Person Pose Estimation	OCHuman	Validation AP	67.8	I²R-Net (1st stage:TransPose-H)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	AP	77.3	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	AP50	91	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	AP75	83.6	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	APL	84.5	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	APM	73	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	COCO (Common Objects in Context)	AR	82.1	I²R-Net (1st stage:HRFormer-B)
1 Image, 2*2 Stitchi	CrowdPose	AP Easy	83.8	I²R-Net (1st stage: HRFormer-B)
1 Image, 2*2 Stitchi	CrowdPose	AP Hard	69.3	I²R-Net (1st stage: HRFormer-B)
1 Image, 2*2 Stitchi	CrowdPose	AP Medium	78.1	I²R-Net (1st stage: HRFormer-B)
1 Image, 2*2 Stitchi	CrowdPose	mAP @0.5:0.95	77.4	I²R-Net (1st stage: HRFormer-B)
1 Image, 2*2 Stitchi	OCHuman	AP50	85	I²R-Net (1st stage:TransPose-H)
1 Image, 2*2 Stitchi	OCHuman	AP75	72.8	I²R-Net (1st stage:TransPose-H)
1 Image, 2*2 Stitchi	OCHuman	Validation AP	67.8	I²R-Net (1st stage:TransPose-H)

I^2R-Net: Intra- and Inter-Human Relation Network for Multi-Person Pose Estimation

Abstract

Results

Related Papers

I^2R-Net: Intra- and Inter-Human Relation Network for Multi-Person Pose Estimation

Abstract

Results

Related Papers