TasksSotADatasetsPapersMethodsSubmitAbout
Papers With Code 2

A community resource for machine learning research: papers, code, benchmarks, and state-of-the-art results.

Explore

Notable BenchmarksAll SotADatasetsPapersMethods

Community

Submit ResultsAbout

Data sourced from the PWC Archive (CC-BY-SA 4.0). Built by the community, for the community.

Papers/Image Projective Transformation Rectification with Synthet...

Image Projective Transformation Rectification with Synthetic Data for Smartphone-captured Chest X-ray Photos Classification

Chak Fong Chong, Yapeng Wang, Benjamin Ng, Wuman Luo, Xu Yang

2022-10-12Image ClassificationMedical Image Classification
PaperPDFCode(official)

Abstract

Classification on smartphone-captured chest X-ray (CXR) photos to detect pathologies is challenging due to the projective transformation caused by the non-ideal camera position. Recently, various rectification methods have been proposed for different photo rectification tasks such as document photos, license plate photos, etc. Unfortunately, we found that none of them is suitable for CXR photos, due to their specific transformation type, image appearance, annotation type, etc. In this paper, we propose an innovative deep learning-based Projective Transformation Rectification Network (PTRN) to automatically rectify CXR photos by predicting the projective transformation matrix. To the best of our knowledge, it is the first work to predict the projective transformation matrix as the learning goal for photo rectification. Additionally, to avoid the expensive collection of natural data, synthetic CXR photos are generated under the consideration of natural perturbations, extra screens, etc. We evaluate the proposed approach in the CheXphoto smartphone-captured CXR photos classification competition hosted by the Stanford University Machine Learning Group, our approach won first place with a huge performance improvement (ours 0.850, second-best 0.762, in AUC). A deeper study demonstrates that the use of PTRN successfully achieves the classification performance on the spatially transformed CXR photos to the same level as on the high-quality digital CXR images, indicating PTRN can eliminate all negative impacts of projective transformation on the CXR photos.

Results

TaskDatasetMetricValueModel
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.906LBC-v2 (ensemble)
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.6LBC-v2 (ensemble)
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.906LBC-v2 (ensemble)
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.6LBC-v2 (ensemble)
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.899LBC-v0 (ensemble)
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.4LBC-v0 (ensemble)
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.899LBC-v0 (ensemble)
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.4LBC-v0 (ensemble)
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.896Stellarium-CheXpert-Local
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.4Stellarium-CheXpert-Local
Multi-Label ClassificationCheXpertAVERAGE AUC ON 14 LABEL0.896Stellarium-CheXpert-Local
Multi-Label ClassificationCheXpertNUM RADS BELOW CURVE1.4Stellarium-CheXpert-Local
ClassificationCheXphotoMean AUC0.85PTRN
Medical Image ClassificationCheXphotoMean AUC0.85PTRN

Related Papers

Automatic Classification and Segmentation of Tunnel Cracks Based on Deep Learning and Visual Explanations2025-07-18Adversarial attacks to image classification systems using evolutionary algorithms2025-07-17Efficient Adaptation of Pre-trained Vision Transformer underpinned by Approximately Orthogonal Fine-Tuning Strategy2025-07-17Federated Learning for Commercial Image Sources2025-07-17MUPAX: Multidimensional Problem Agnostic eXplainable AI2025-07-17Hashed Watermark as a Filter: Defeating Forging and Overwriting Attacks in Weight-based Neural Network Watermarking2025-07-15Transferring Styles for Reduced Texture Bias and Improved Robustness in Semantic Segmentation Networks2025-07-14FedGSCA: Medical Federated Learning with Global Sample Selector and Client Adaptive Adjuster under Label Noise2025-07-13