U.S., July 30 -- ClinicalTrials.gov registry received information related to the study (NCT07732985) titled 'Comparative Analysis of Diagnostic Accuracy and Case Difficulty Assessment of Three Large Language Models' on July 10.
Brief Summary: This prospective, blinded diagnostic accuracy study aims to compare the performance of three large language models-ChatGPT (GPT-5.5 Pro), Gemini 3.1 Pro, and Claude Opus 4.7-in endodontic diagnosis and case difficulty assessment. The models will be evaluated against expert consensus as the reference standard using standardized clinical data and periapical radiographs. Diagnostic accuracy, sensitivity, specificity, and agreement with expert consensus will be assessed to determine the potential of LLMs ...