PubMed چکیده/رکورد

Comparison of Two AI Chatbots for Diagnosis and Providing Treatment Suggestions in Retinopathy of Prematurity: Retrospective Study.

استودیوی صوتی مقاله

پخش حرفه‌ای فارسی و انگلیسی

در حال بررسی نسخه‌های صوتی ذخیره‌شده…

صوت تولیدشده با هوش مصنوعی است. برای کاربرد علمی یا درمانی، متن و منبع اصلی را بررسی کنید.
خواندن هوشمند فارسی و انگلیسی در حال آماده‌سازی صداهای مرورگر…
تنظیم صدای طبیعی و سرعت

صداهایی که در نامشان «Natural»، «Neural» یا «Online» دیده می‌شود معمولاً طبیعی‌ترند. انتخاب صدا به صداهای نصب‌شده در ویندوز و مرورگر شما بستگی دارد.

چکیده اصلی

BACKGROUND: Retinopathy of Prematurity (ROP) is a leading cause of preventable childhood blindness; yet, a global shortage of experienced pediatric ophthalmologists impedes timely diagnosis and treatment. While emerging AI chatbots are promising clinical decision-support tools in some ophthalmic diseases, their performance in ROP diagnosis and providing treatment suggestions remains uncertain. OBJECTIVE: This study aimed to compare the performance of Google's Gemini 2.5 Pro and OpenAI's ChatGPT o4-mini in ROP diagnosis and providing treatment suggestions against the gold standard of clinical consensus. METHODS: A retrospective analysis was conducted on 70 infants (140 eyes) with treatment-requiring ROP, each providing structured clinical text data and wide-field fundus images. We adopted a 2-stage prompting strategy for AI chatbots, instructing them first to generate ROP diagnoses (including zone, stage, and presence of plus disease), and subsequently to provide treatment suggestions. After collecting the generated responses, we assessed their performance by comparing the consistency of their diagnosis and treatment suggestions with the consensus gold standard. Furthermore, 2 independent specialists quantitatively assessed the outputs of Gemini 2.5 Pro and ChatGPT o4-mini using the ROP-specific Global Quality Score (GQS), which is a 5-point scale ranging from 1 (unusable) to 5 (excellent). Statistical significance was set at P<.05, and all statistical analyses were performed using R (version 4.4.1; R Foundation for Statistical Computing). RESULTS: For the tasks of ROP zoning and staging, the consistency rates between Gemini 2.5 Pro and ChatGPT o4-mini were 79.3% (111/140) vs 85.7% (120/140; zone), 64.3% (90/140) vs 70.0% (98/140; stage), respectively. For the task of treatment requirement, the rates (also referred to as sensitivity) were 93.6% (131/140) vs 90.7% (127/140), respectively. None of these differences were statistically significant (P>.05). However, Gemini 2.5 Pro showed significantly better performance in plus disease identification (consistency: 108/140, 77.1% vs 80/140, 57.1%; P=.006), while ChatGPT o4-mini demonstrated significantly higher guideline adherence in treatment modality suggestions based on gold-standard ROP diagnoses (consistency: 91/140, 65.0%; vs 55/140, 39.3%; P=.01). Compared to Gemini 2.5 Pro, ChatGPT o4-mini performed better in providing ROP treatment suggestions (GQS score; P=.001), while the 2 AI chatbots had comparable GQS scores in diagnostic tasks. CONCLUSIONS: ChatGPT o4-mini shows greater promise in generating evidence-based treatment suggestions based on gold-standard diagnoses, whereas Gemini 2.5 Pro shows advantages in visual interpretation, supporting its potential for targeted ROP diagnostic screening, particularly in identifying plus disease. As these AI chatbots continue to evolve, their performance merits further validation using larger cohorts.

متن کامل اصلی

متن در JumpToDate ذخیره نشده است.

برای بررسی دسترسی کتابخانه‌ای یا خرید، رکورد اصلی را باز کنید.

رفتن به منبع اصلی

کلیدواژه‌ها

Artificial intelligence chatbotdiagnosisretinopathy of prematuritytreatment suggestionwide-field fundus image
در همین زیرشاخه

مقاله‌های مرتبط

PubMed2028

Association between mid-to-late pregnancy gestational weight gain and adverse birth outcomes among women with gestational diabetes mellitus.

BACKGROUND AND OBJECTIVES: The gestational weight gain (GWG) range during mid-to-late pregnancy associated with the lowest combined risk of adverse birth outcomes in women with gestational diabetes mellitus (GDM) remains unclear. This study aims to examine the associations between GWG and adverse birth out-comes among women with GDM. METHODS AND STUDY DESIGN: This study included a cohort of 1,673 pregnant women with GDM. GWG was define…

PubMed2026

[THE CHARACTERISTICS OF COURSE OF PREGNANCY RESULTED FROM EXTRA-CORPOREAL IMPREGNATION AGAINST THE BACKGROUND OF GESTATIONAL DIABETES MELLITUS DEPENDING ON TERMS OF MANIFESTATION].

The pregnancy that occurred due to in vitro fertilization is characterized by higher risk of development of gestational diabetes mellitus. The purpose of the study was to investigate course and outcomes of pregnancy resulted from in vitro fertilization in women with gestational diabetes mellitus, depending on time of its manifestation. The analysis of course of pregnancy, childbirth and condition of newborns in 179 women with gestation…

PubMed2026

Therapeutic drug monitoring-guided individualized caffeine dosing for apnea of prematurity: Clinical efficacy and association with early neurobehavioral outcomes in preterm infants.

BACKGROUND: Given the high incidence of apnea of prematurity (AOP) and the repeated hypoxia-induced nerve damage, treatment optimization from the dosing perspective is critical. Caffeine is the current first-line therapeutic drug for AOP. However, the conventional dose results in low blood concentration compliance, with significant variation among individuals. OBJECTIVES: To explore the clinical efficacy and safety of a therapeutic dru…

PubMed2026

A National Pediatric Cohort Study on In-Hospital Cardiac Arrest in Sweden.

BACKGROUND: National data on pediatric in-hospital cardiac arrest (pIHCA) are limited, and cohorts from highly specialized pediatric centers may not reflect the broader hospital population. We aimed to describe pIHCA reported across Swedish hospitals, with particular focus on hospitals outside the two national centers for highly specialized pediatric cardiac care. METHODS: This retrospective observational registry study included patien…