PubMed دسترسی آزاد

Artificial intelligence meets pediatric orthopedics: A comparative analysis of ChatGPT-4o, Gemini 2.0, and Claude 3.5 in detecting supracondylar humeral fractures.

استودیوی صوتی مقاله

پخش حرفه‌ای فارسی و انگلیسی

در حال بررسی نسخه‌های صوتی ذخیره‌شده…

صوت تولیدشده با هوش مصنوعی است. برای کاربرد علمی یا درمانی، متن و منبع اصلی را بررسی کنید.
خواندن هوشمند فارسی و انگلیسی در حال آماده‌سازی صداهای مرورگر…
تنظیم صدای طبیعی و سرعت

صداهایی که در نامشان «Natural»، «Neural» یا «Online» دیده می‌شود معمولاً طبیعی‌ترند. انتخاب صدا به صداهای نصب‌شده در ویندوز و مرورگر شما بستگی دارد.

چکیده اصلی

BACKGROUND: Supracondylar humeral fractures constitute 10-16% of pediatric skeletal injuries, requiring timely diagnosis to prevent neurovascular complications. Developmental variations in pediatric bone structures pose diagnostic challenges for clinicians. This study evaluated three next-generation large language models (LLMs) (ChatGPT-4o, Gemini 2.0, Claude 3.5) for detecting pediatric supracondylar humeral fractures and their classification according to the Gartland system. METHODS: This retrospective observational study included 300 pediatric patients (150 with supracondylar humeral fractures confirmed by expert consensus, 150 without fractures) aged 2-10 years presenting to the Emergency Department of the Bilkent City Hospital (October 2022-January 2025). Two-view elbow radiographs were presented to each LLM three times on different days. Diagnostic accuracy was evaluated using overall accuracy (all three responses correct), strict accuracy (≥2 correct responses), and ideal accuracy (≥1 correct response). Response consistency was assessed using Fleiss' Kappa coefficient. Fractures were classified according to modified Gartland criteria. RESULTS: Gemini 2.0 demonstrated highest sensitivity (68.4%) followed by Claude 3.5 (58.7%) and ChatGPT-4o (19.3%) for fracture detection (p < 0.001). Ideal accuracy rates were 83.3%, 78.7%, and 27.3% respectively. Although ideal accuracy rates exceeded 91% in non-fracture cases, specificity remained low (33.1-36.0%), indicating a high rate of false-positive classifications. Response consistency was very good for ChatGPT-4o (κ = 0.69) and Gemini 2.0 (κ = 0.61), good for Claude 3.5 (κ = 0.44). For Gartland classification, Gemini 2.0 achieved highest accuracy: Type I (83.3%), Type II (62.4%), Type III (68.7%). CONCLUSION: Current LLMs demonstrate limited capability as independent diagnostic tools for pediatric supracondylar humeral fractures. Gemini 2.0's 68.4% sensitivity indicates these technologies require specialized pediatric training before clinical implementation. However, their potential as assistive tools for triage and assessment warrants further development of pediatric-specific models.

متن کامل اصلی

نسخه دارای مجوز در منبع علمی در دسترس است.

لینک مستقیم از metadata منبع گرفته شده و در تب جدید باز می‌شود.

باز کردن متن کامل
در همین زیرشاخه

مقاله‌های مرتبط

PubMed2026

Sustainable metallic biomaterials for orthopaedic implants: a comprehensive review of biodegradable and conventional metals.

The selection of biomaterial is crucial for the long-term success of implants. Materials that perform an adequate function and reduce negative biological responses should be taken. Due to their good mechanical strength, stainless steel, titanium, and Co-based alloys have been utilized for implant purposes; however, their permanent nature and very low corrosion rates may lead to long-term clinical complications. Researchers are looking …

PubMed2026

Leaving orthopaedic surgical training: the LOST surgeons - a qualitative study exploring why UK trauma and orthopaedic registrars discontinue surgical training.

OBJECTIVES: The study aimed to explore why trauma and orthopaedic registrars decide to discontinue surgical training. Understanding the factors that influence a decision to leave may help to inform changes which enhance the experiences of surgeons and retention of the future workforce. DESIGN: Qualitative study using semi-structured interviews. SETTING: Between October 2022 and March 2026, interviews were conducted with participants wh…

PubMed2026

e-Learning, Distance Education, and Virtual and Augmented Reality in Orthopedic Training: European Cross-Sectional Survey of Trainee Acceptance Guided by the Technology Acceptance Model and Unified Theory of Acceptance and Use of Technology.

BACKGROUND: Digital technologies increasingly shape postgraduate medical education, yet orthopedic and trauma training face unique challenges because of the tactile, procedurally focused skills involved. Digital tools partially address these needs, but gaps remain, particularly across diverse European contexts. OBJECTIVE: Our primary aim was to quantitatively assess predictors of digital learning technology acceptance (e-learning, dist…

PubMed2026

Shoulder Procedure Volumes in Orthopaedic Residency: Long-Term Disparities and a Case for Arthroplasty Minimums.

BACKGROUND: Total shoulder arthroplasty (TSA) has experienced rapid growth. Yet graduating orthopaedic surgery resident (GOSR) TSA volumes are not individually reported due to its exclusion from the Accreditation Council for Graduate Medical Education (ACGME) case minimum list. Rather, TSAs are grouped within the shoulder repair/revision/reconstruction (RRR) category. This study evaluated long-term trends and identified disparities in …