Library
PubMed Central Open Access
research article
Professional
Open access

Tensor enhanced chest cancer classification via CNN and Vision Transformer models

Source: PubMed Central Open Access, NCBI / U.S. National Library of Medicine

PLOS OneLast synced 6/4/2026Status: syncedPMID: 42228733 pmidDOI: 10.1371/journal.pone.0348863

Lung diseases, particularly lung cancer, remain a leading cause of mortality worldwide, accounting for approximately 1.8 million deaths annually. Early and accurate diagnosis is critical for improving patient outcomes. This study also introduces a unified platform for evaluating multiple convolutional neural network architectures and comparing them to a Vision Transformer model while utilizing a common tensor-based preprocessing pipeline for classifying lung cancer with CT/PET-CT imaging. To enhance model adaptability, all input images were initially converted into tensors prior to training, enabling implicit fine-tuning without altering the original architecture. The YOLOTransfer dataset, comprising diverse and annotated medical images, was used to benchmark model performance. Classical CNN models such as AlexNet, VGG-16, ResNet-50, DenseNet, and EfficientNet were compared against ViT in terms of accuracy, sensitivity, specificity, F1-score, and AUC-ROC. Among all models, ResNet-50 and EfficientNet achieved the highest accuracy, while the Vision Transformer showed competitive results in capturing complex global patterns. The findings highlight the complementary strengths of convolutional and transformer-based architectures for medical image analysis and demonstrate the feasibility of deep learning approaches for lung cancer detection.

Educational only
This information is for general education and is not medical advice. Always talk to a licensed U.S. clinician about your situation, medications, or treatment decisions.