Back to results
Bibliographic record · Consultation and access
Artículo

A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes

Yin Zhou et al · Elsevier · 2026

Supplementary material available
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.

Resource access

Open the content from the main option or choose another available source.

DOAJ DOAJ Articles
Entrar por DOAJ
Main access

Supplementary material available

El enlace apunta a material asociado, anexos, tablas, datos o página complementaria. No se marca como libro/texto completo.
Open material

Summary

Descripción general del contenido del recurso.

Background: Early-stage high-grade lung invasive adenocarcinoma (IAC) has poor prognosis and is hard to identify using conventional radiological assessment. Current reliance on postoperative histology to identify high-grade subtypes delays risk-adapted surgical planning. Three-dimensional (3D) vision transformers (ViTs) may improve prediction by modeling long-range dependencies in computed tomography (CT) scans. We aimed to develop and validate 3D-ViT and Swin Transformer (SwinT) for preoperative CT-based prediction of early-stage high-grade IAC subtypes (micropapillary/solid), benchmarking against ResNet. Methods: A multicenter cohort of 1028 patients with surgically confirmed early-stage lung adenocarcinoma was divided into training (n = 806), validation (n = 100), and external test (n = 122) sets. 3D-ViT, SwinT, and ResNet models were trained on CT to classify nodules harboring high-grade histologic patterns. A novel decision-aid tool for IAC surgery was provided. Performance was evaluated using area under the curve (AUC), accuracy, sensitivity, specificity, and precision. Attention mapping was performed to interpret 3D-ViT decision-making. Results: The 3D-ViT model achieved AUC values of 0.856 (95% CI: 0.845–0.877) (validation) and 0.806 (95% CI: 0.790–0.816) (testing), compared to 0.854 (95% CI: 0.841–0.872) (validation) and 0.760 (95% CI: 0.743–0.776) (testing) for the ResNet baseline. 3D-ViT showed balanced accuracy, sensitivity, specificity, and precision in the validation set. In external testing, 3D-ViT significantly outperformed ResNet in all metrics with P < 0.01. The SwinT-based AlignSen model from decision-aid tool prioritized sensitivity for high-grade IAC (88.0% validation, 93.4% testing), while maintaining specificity (76.0% validation, 54.1% testing) which significantly outperformed ViT and ResNet-based AlignSen models’ specificity (60.0% and 54% validation, 52.5% and 24.6% testing, respectively). Attention maps highlighted 3D nodule heterogeneity and peripheral irregularities. Conclusion: The 3D-ViT model demonstrated robust accuracy and generalizability in predicting high-grade IAC subtypes using CT images. Integration of preoperative SwinT into clinical workflows may offer a viable alternative to intraoperative pathology subtyping, potentially reducing reliance on frozen sections while optimizing surgical planning.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, Y. Z. E. (2026). A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes. https://doi.org/10.1016/j.imed.2025.11.001

MLA

al, Yin Zhou et. "A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes." 2026. https://doi.org/10.1016/j.imed.2025.11.001.

Chicago

al, Yin Zhou et. 2026. "A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes.". https://doi.org/10.1016/j.imed.2025.11.001.

Harvard

al, Y. Z. E. 2026, A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes, Elsevier, available at: https://doi.org/10.1016/j.imed.2025.11.001 [Accessed 7 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
A pulmonary nodule is worth 8 × 8 × 8 words: Computed tomography-based 3D vision transformer predicts early-stage high-grade lung adenocarcinoma of micropapillary and/or solid subtypes
Author / contributors
Yin Zhou et al
Publisher
Elsevier
Publication year
2026
ISSN
2667-1026
ISSN
2667-1026
Language
English

Subjects

Explore related resources through these subjects.

Copied