Back to results
Bibliographic record · Consultation and access
Artículo de revista

Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study

Hayriye Bektaş Aksoy et al · Frontiers Media S.A · 2026

Supplementary material available
Quick overview. Review the resource’s basic details, then access the content using the main button. This page shows only the information needed to identify, cite, and open the work.
Serial publication

A combined model of BCVA, TRAb, and NLR predicts response to intravenous methylprednisolone in dysthyroid optic neuropathy

This serial publication contains 132 related contents.

Resource access

Open the content from the main option or choose another available source.

DOAJ DOAJ Articles
Entrar por DOAJ
Main access

Supplementary material available

El enlace apunta a material asociado, anexos, tablas, datos o página complementaria. No se marca como libro/texto completo.
Open material

Summary

Descripción general del contenido del recurso.

IntroductionLarge language models (LLMs) may support clinical reasoning, yet real-world outpatient studies integrating structured clinical data with chest radiographs (CXRs) remain limited. We compared three LLMs for pulmonary differential diagnosis in routine clinical practice.MethodsIn this prospective, single-center observational study, consecutive adult outpatients presenting with respiratory complaints between 06 October and 31 December 2025 were enrolled. For each case, a standardized structured clinical form and de-identified CXRs were provided to three LLMs (ChatGPT-5.2, Google Gemini 3 Flash, and Microsoft Copilot) and to three blinded pulmonologists. The primary diagnosis was assigned by the examining pulmonologist, and the reference diagnosis was defined by agreement of at least two blinded pulmonologists. Concordance and Cohen’s kappa were assessed.ResultsA total of 120 patients were included. Agreement among the blinded pulmonologists was high, and agreement between the primary and reference diagnoses was excellent. Compared with the reference diagnosis, ChatGPT-5.2 and Microsoft Copilot showed higher concordance than Google Gemini 3 Flash, with both demonstrating moderate overall agreement. Concordance did not differ by age or sex. Across diagnostic categories, performance was highest for pneumonia/upper respiratory tract infection and asthma.DiscussionIn this real-world pulmonology outpatient cohort, ChatGPT-5.2 and Microsoft Copilot showed better diagnostic concordance than Google Gemini 3 Flash when structured clinical data and CXRs were evaluated together. These findings support the potential role of LLMs as adjunctive decision-support tools in pulmonology, while also indicating that performance remains diagnosis-dependent and insufficient to replace expert clinical judgment.

How to cite

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

al, H. B. A. E. (2026). Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study. https://doi.org/10.3389/fmed.2026.1821639

MLA

al, Hayriye Bektaş Aksoy et. "Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study." 2026. https://doi.org/10.3389/fmed.2026.1821639.

Chicago

al, Hayriye Bektaş Aksoy et. 2026. "Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study.". https://doi.org/10.3389/fmed.2026.1821639.

Harvard

al, H. B. A. E. 2026, Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study, Frontiers Media S.A, available at: https://doi.org/10.3389/fmed.2026.1821639 [Accessed 9 Aug. 2026].

Share and print

Save the record, copy its permanent link, or print it as a PDF.

Export reference

You can export the record in common formats for use in a reference manager.

Resource details

Bibliographic information to help confirm that this is the correct material.

Title
Evaluation of large language models in a pulmonology outpatient clinic using structured clinical data and chest radiographs: a single-center prospective observational study
Author / contributors
Hayriye Bektaş Aksoy et al
Publisher
Frontiers Media S.A
Publication year
2026
ISSN
2296-858X
ISSN
2296-858X
Language
English

Subjects

Explore related resources through these subjects.

Copied