Explainable AI for Image-Based Auxiliary Assessment of Radiation Pneumonitis on Post-Treatment CT Images Using GAN Grad-CAM and Ensemble Learning Techniques
Tsair-Fwu Lee, Guang-Zhi Lin, Wen-Ping Yun, Yi-Lun Liao, Liyun Chang, Chao-Hong Liu, Cheng-Shie Wuu, Yu-Chang Hu, Yu-Wei Lin, Pei-Ju Chao, Yang-Wei HsiehObjective: This study aims to develop a preliminary explainable image-based auxiliary assessment framework for radiation pneumonitis (RP) by integrating a generative adversarial network (GAN), Gradient-weighted Class Activation Mapping (Grad-CAM), and ensemble learning to improve image-based classification performance, interpretability, and computational efficiency. Methods: Chest CT images from 46 lung cancer patients who underwent VMAT were retrospectively collected, yielding 542 RP and 1857 non-RP images. Images were preprocessed using Otsu-based segmentation and standardized before being input into the RP-GAN for feature extraction. Grad-CAM was applied to qualitatively visualize attention patterns across convolutional layers and guide feature-layer selection, while PCA reduced the extracted features from 12,288 to 730 dimensions. An ensemble stacking classifier combining RF, SVM, KNN, and XGBoost with logistic regression as a meta-learner was constructed. Model performance was evaluated using AUC, accuracy, PPV, NPV, specificity, recall, and F1-score. Results: Grad-CAM highlighted conv2d_4 and conv2d_5 as the most informative layers. PCA reduced training time from 49 min to 49 s with minimal performance loss. In the internal hold-out test set, the ensemble model achieved the highest point estimates for AUC and accuracy among the evaluated classifiers (AUC = 0.921; accuracy = 87.1%). The model identified RP-positive CT slices and provided qualitative visual cues for suspected RP-related regions. Conclusions: The proposed GAN–Grad-CAM ensemble framework showed promising internal classification performance for post-treatment CT-based RP image assessment with substantially improved computational efficiency. Its qualitative visual outputs may support clinical image review, although further external validation and quantitative localization assessment are required before clinical application as an auxiliary image-review tool.