PubMed دسترسی آزاد

Classification Performance of General-Purpose Multimodal Large Language Models Across Orthodontic Radiographic Tasks: A Comparative Study of ChatGPT, Gemini, and Claude.

استودیوی صوتی مقاله

پخش حرفه‌ای فارسی و انگلیسی

در حال بررسی نسخه‌های صوتی ذخیره‌شده…

صوت تولیدشده با هوش مصنوعی است. برای کاربرد علمی یا درمانی، متن و منبع اصلی را بررسی کنید.
خواندن هوشمند فارسی و انگلیسی در حال آماده‌سازی صداهای مرورگر…
تنظیم صدای طبیعی و سرعت

صداهایی که در نامشان «Natural»، «Neural» یا «Online» دیده می‌شود معمولاً طبیعی‌ترند. انتخاب صدا به صداهای نصب‌شده در ویندوز و مرورگر شما بستگی دارد.

چکیده اصلی

Background and Objectives: General-purpose multimodal large language models (MLLMs) can interpret radiographic images, but their classification performance across orthodontic tasks remains uncertain. This study compared the classification performance of ChatGPT, Gemini, and Claude on lateral cephalometric, hand-wrist, and panoramic radiographs. Materials and Methods: This retrospective diagnostic accuracy study included 250 individuals, each contributing one lateral cephalometric, hand-wrist, and panoramic pretreatment radiograph (750 total). Reference classifications were established by two experienced orthodontists, with disagreements resolved by consensus. Lateral cephalometric radiographs were classified as skeletal Class I, II, or III based on the ANB angle according to Steiner analysis; hand-wrist radiographs as prepubertal, pubertal, or postpubertal; and panoramic radiographs as early mixed, late mixed, or permanent dentition. Each image was evaluated once by each AI platform using identical Turkish prompts in separate chat sessions. Classification accuracy, balanced accuracy, macro-F1, class-specific metrics, and reference agreement were assessed. Generalized estimating equations (GEE) assessed platform, radiograph type, and interaction effects on correct classification. Results: ChatGPT had the highest hand-wrist accuracy (81.6%; 95% CI, 76.3-85.9), whereas Gemini had the highest panoramic accuracy (92.8%; 95% CI, 88.9-95.4). Lateral cephalometric accuracies were 70.0%, 64.4%, and 64.8% for ChatGPT, Gemini, and Claude, respectively, with no significant interplatform difference (p = 0.336). The platform × radiograph type interaction was significant (Wald χ2 = 42.52; df = 4; p < 0.001). Agreement with the reference standard was highest for ChatGPT on hand-wrist radiographs (κw = 0.758) and Gemini on panoramic radiographs (κw = 0.883). Conclusions: Classification performance was task- and platform-dependent, with no model consistently achieving the highest performance. For the predefined classification tasks, these models should not be used as standalone tools for these classification tasks. Their potential as decision-support tools requires prospective evaluation of AI-assisted clinician performance and external validation.

متن کامل اصلی

نسخه دارای مجوز در منبع علمی در دسترس است.

لینک مستقیم از metadata منبع گرفته شده و در تب جدید باز می‌شود.

باز کردن متن کامل

کلیدواژه‌ها

age determination by skeletonartificial intelligencecephalometrylarge language modelsorthodonticspanoramicradiography
در همین زیرشاخه

مقاله‌های مرتبط

PubMed2026

Torque expression in directly printed and thermoformed clear aligners: a controlled in vitro biomechanical comparison of power ridge effects.

BACKGROUND: Torque control of anterior teeth remains a major challenge in clear aligner therapy. Power ridges have been introduced to improve buccolingual root control; however, their biomechanical performance in directly printed aligners remains insufficiently investigated. The aim of this study was to compare the torque efficiency of power ridges in thermoformed and directly printed aligners. METHODOLOGY: A controlled in-vitro study …

PubMed2026

Predicting orthodontic appointment adherence in adolescents and parents using an extended behavioral model.

BACKGROUND: Poor adherence in orthodontic care is a complex issue involving both children and their parents. The study aimed to expand the theory of planned behavior (TPB) by integrating sense of coherence (SOC) to predict appointment attendance in adolescent orthodontic patients. MATERIALS AND METHODS: This prospective cohort study recruited adolescents from the University of Alberta Graduate Orthodontic Clinic. Baseline questionnaire…

PubMed2026

Surgical and orthodontic management of class III malocclusion in monozygotic twins using bilateral sagittal split osteotomy.

This case report discusses the genetic influence and treatment of class III malocclusion in identical twin sisters, in their early 20s, who presented with mandibular prognathism. Both exhibited concave profiles with hypodivergent growth patterns, but Sister A had a more pronounced concavity. Diagnosis revealed skeletal class III malocclusion with mandibular prognathism and negative overjet (9 mm in Sister A, 6 mm in Sister B). Treatmen…

PubMed2026

H3K18la in periodontal ligament fibroblasts regulates immune niches and alveolar bone remodeling under mechanical loads.

The metabolism and remodeling of alveolar bone are highly active, which dynamically respond to mechanical force. However, the mechanism of how alveolar bone responds to mechanical force remains elusive. Orthodontics is essentially a physiological process of mechanical force-driven alveolar bone remodeling. Here, we demonstrated that orthodontic tension promoted the expression of glycolysis-related genes in periodontal ligament (PDL) fi…