PubMed چکیده/رکورد

Comparing Speed and Accuracy of Artificial Intelligence Large Language Models on the Orthopedic In-Training Examination.

استودیوی صوتی مقاله

پخش حرفه‌ای فارسی و انگلیسی

در حال بررسی نسخه‌های صوتی ذخیره‌شده…

صوت تولیدشده با هوش مصنوعی است. برای کاربرد علمی یا درمانی، متن و منبع اصلی را بررسی کنید.
خواندن هوشمند فارسی و انگلیسی در حال آماده‌سازی صداهای مرورگر…
تنظیم صدای طبیعی و سرعت

صداهایی که در نامشان «Natural»، «Neural» یا «Online» دیده می‌شود معمولاً طبیعی‌ترند. انتخاب صدا به صداهای نصب‌شده در ویندوز و مرورگر شما بستگی دارد.

چکیده اصلی

OBJECTIVES: Large language models (LLMs), such as Open AI's Chat Generative Pre-Trained Transformer (GPT)-4 and Google Gemini, have gained significant attention for their ability to process complex language patterns and are being used increasingly in fields such as medicine, where they assist in learning, collaboration, and patient care. Although prior studies have evaluated LLMs on medical licensing examinations, limited research compares their performance on orthopedic-specific assessments. This study aims to assess the accuracy and response speed of ChatGPT-3.5, ChatGPT-4, Microsoft Copilot, and Gemini on the Orthopedic In-Training Examination (OITE). METHODS: Questions from the 2020-2022 OITE were extracted from the American Academy of Orthopaedic Surgeons' question bank. Each question, along with four answer choices, was manually input into the LLMs without response prompts or feedback. Response accuracy and speed were recorded, with timing measured from the moment the question was submitted until an answer was generated. RESULTS: Out of 1582 prompts, ChatGPT-4 demonstrated the highest accuracy (67.09%±0.08%), significantly outperforming ChatGPT-3.5, Microsoft Copilot, and Gemini (P<0.001). ChatGPT-3.5 was the fastest, with an average response time of 5.41±0.10 seconds. Both ChatGPT-3.5 and ChatGPT-4 responded significantly faster than Gemini and Microsoft Copilot (P<0.001). CONCLUSIONS: ChatGPT-4 exhibited the highest accuracy on OITE questions, and ChatGPT-3.5 was the fastest. Gemini and Copilot were generally less accurate in their responses and had a slower response time. These findings highlight the potential of LLMs in orthopedic education and emphasize the need for further research to explore their broader applications in medical training and decision making.

متن کامل اصلی

متن در JumpToDate ذخیره نشده است.

برای بررسی دسترسی کتابخانه‌ای یا خرید، رکورد اصلی را باز کنید.

رفتن به منبع اصلی
در همین زیرشاخه

مقاله‌های مرتبط

PubMed2026

Sustainable metallic biomaterials for orthopaedic implants: a comprehensive review of biodegradable and conventional metals.

The selection of biomaterial is crucial for the long-term success of implants. Materials that perform an adequate function and reduce negative biological responses should be taken. Due to their good mechanical strength, stainless steel, titanium, and Co-based alloys have been utilized for implant purposes; however, their permanent nature and very low corrosion rates may lead to long-term clinical complications. Researchers are looking …

PubMed2026

Artificial intelligence meets pediatric orthopedics: A comparative analysis of ChatGPT-4o, Gemini 2.0, and Claude 3.5 in detecting supracondylar humeral fractures.

BACKGROUND: Supracondylar humeral fractures constitute 10-16% of pediatric skeletal injuries, requiring timely diagnosis to prevent neurovascular complications. Developmental variations in pediatric bone structures pose diagnostic challenges for clinicians. This study evaluated three next-generation large language models (LLMs) (ChatGPT-4o, Gemini 2.0, Claude 3.5) for detecting pediatric supracondylar humeral fractures and their classi…

PubMed2026

Leaving orthopaedic surgical training: the LOST surgeons - a qualitative study exploring why UK trauma and orthopaedic registrars discontinue surgical training.

OBJECTIVES: The study aimed to explore why trauma and orthopaedic registrars decide to discontinue surgical training. Understanding the factors that influence a decision to leave may help to inform changes which enhance the experiences of surgeons and retention of the future workforce. DESIGN: Qualitative study using semi-structured interviews. SETTING: Between October 2022 and March 2026, interviews were conducted with participants wh…

PubMed2026

e-Learning, Distance Education, and Virtual and Augmented Reality in Orthopedic Training: European Cross-Sectional Survey of Trainee Acceptance Guided by the Technology Acceptance Model and Unified Theory of Acceptance and Use of Technology.

BACKGROUND: Digital technologies increasingly shape postgraduate medical education, yet orthopedic and trauma training face unique challenges because of the tactile, procedurally focused skills involved. Digital tools partially address these needs, but gaps remain, particularly across diverse European contexts. OBJECTIVE: Our primary aim was to quantitatively assess predictors of digital learning technology acceptance (e-learning, dist…