Back to results
Bibliographic record · Consultation and access
Artículo

Bridging radiology and pathology: domain-generalized cross-modal learning for clinical

Xiang Zhong et al · Nature Portfolio · 2026

Supplementary material available
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.

Resource access

Open the content from the main option or choose another available source.

DOAJ DOAJ Articles
Entrar por DOAJ
Main access

Supplementary material available

El enlace apunta a material asociado, anexos, tablas, datos o página complementaria. No se marca como libro/texto completo.
Open material

Summary

Descripción general del contenido del recurso.

Abstract Reliable interpretation of clinical imaging requires integrating complementary evidence across modalities, yet most AI systems remain limited by single-modality analysis and poor generalization across institutions. We propose a unified cross-modal framework that bridges mammography and histopathology for breast cancer diagnosis through: (1) a shared vision transformer encoder with lightweight modality-specific adapters, (2) a weakly supervised patient-level contrastive alignment module that learns cross-modal correspondences without pixel-level supervision, (3) domain generalization strategies combining MixStyle augmentation and invariant risk minimization, and (4) causal test-time adaptation for unseen target domains. The model jointly addresses classification, lesion localization, and pathological grading while generating reasoning-guided attention maps that explicitly link suspicious mammographic regions with corresponding histopathological evidence. Evaluated on four public benchmarks (CBIS-DDSM, INbreast, BACH, CAMELYON16/17), the framework consistently outperforms state-of-the-art unimodal, multimodal, and domain generalization baselines, achieving mean AUC of 0.90 under rigorous leave-one-domain-out evaluation and substantially smaller domain gaps (0.03 vs. 0.06–0.10). Visualization and interpretability analyses further confirm that predictions align with clinically meaningful features, supporting transparency and trust. By advancing multimodal integration, cross-institutional robustness, and explainability, this study represents a step toward clinically deployable AI systems for diagnostic decision support.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, X. Z. E. (2026). Bridging radiology and pathology: domain-generalized cross-modal learning for clinical. https://doi.org/10.1038/s41746-026-02423-w

MLA

al, Xiang Zhong et. "Bridging radiology and pathology: domain-generalized cross-modal learning for clinical." 2026. https://doi.org/10.1038/s41746-026-02423-w.

Chicago

al, Xiang Zhong et. 2026. "Bridging radiology and pathology: domain-generalized cross-modal learning for clinical.". https://doi.org/10.1038/s41746-026-02423-w.

Harvard

al, X. Z. E. 2026, Bridging radiology and pathology: domain-generalized cross-modal learning for clinical, Nature Portfolio, available at: https://doi.org/10.1038/s41746-026-02423-w [Accessed 9 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Bridging radiology and pathology: domain-generalized cross-modal learning for clinical
Author / contributors
Xiang Zhong et al
Publisher
Nature Portfolio
Publication year
2026
ISSN
2398-6352
ISSN
2398-6352
Language
English
Copied