Comparative Efficacy of Trained Versus Untrained Generative Artificial Intelligence Platforms in Providing Case-Based Multiple-Choice Questions on Traumatic Dental Injuries in the Pediatric Dentistry Curriculum: A Cross-Sectional Study.

Comparative Efficacy of Trained Versus Untrained Generative Artificial Intelligence Platforms in Providing Case-Based Multiple-Choice Questions on Traumatic Dental Injuries in the Pediatric Dentistry Curriculum: A Cross-Sectional Study.

Puranik, Chaitanya P; Turley, Jakob; Pickett-Nairne, Kaci; Katebzadeh, Shahbaz
Dental traumatology : official publication of International Association for Dental Traumatology 2026
5
puranik2026comparative

Abstract

To determine the comparative efficacy of trained versus untrained generative artificial intelligence platforms in providing multiple-choice questions on traumatic dental injuries in a pediatric dentistry curriculum. In this cross-sectional study, a standardized prompt was used on three generative artificial intelligence platforms, accessed via web interfaces in 2025 (United States), to generate case-based multiple-choice questions on four domains: avulsion, crown-root fractures, primary teeth, and permanent teeth injuries. The three generative artificial intelligence platforms used were: ScholarGPT, untrained, and trained ChatGPT4o (OpenAI). Evidence-based guidelines from the International Association for Dental Traumatology were used to train the platform. The generative artificial intelligence platforms were asked to select one correct answer from four choices and provide a rationale for the selection. Two calibrated, masked, board-certified pediatric dentists scored the case-based multiple-choice questions using a validated Artificial Intelligence Study Material Assessment and Reliability tool and noted subjective responses for the questions. Statistical analyses were performed with an alpha value of 0.05. All three generative artificial intelligence platforms demonstrated no statistically significant difference (p > 0.05) in terms of their Artificial Intelligence Study Material Assessment and Reliability scores. Some questions had incomplete clinical information (37%), while the options were simplistic (56%) with incorrect rationale (31%). Newer or trained generative artificial intelligence platforms have scores similar to those of untrained platforms, suggesting that publicly available evidence-based information from the International Association for Dental Traumatology enabled the platforms to access these resources for enhanced accuracy. The newer or trained generative artificial intelligence platforms show potential to augment dental trauma education within pediatric dentistry through the development of case-based multiple-choice questions; however, limitations in rationale accuracy and answer quality highlight the need for expert oversight.

Citation

ID: 15225
Ref Key: puranik2026comparative
Use this key to autocite in SciMatic or Thesis Manager

References

Blockchain Verification

Account:
NFT Contract Address:
0x95644003c57E6F55A65596E3D9Eac6813e3566dA
Article ID:
15225
Unique Identifier:
10.1111/edt.70095
Network:
Scimatic Chain (ID: 481)
Loading...
Blockchain Readiness Checklist
Authors
Abstract
Journal Name
Year
Title
5/5
Creates 1,000,000 NFT tokens for this article
Token Features:
  • ERC-1155 Standard NFT
  • 1 Million Supply per Article
  • Transferable via MetaMask
  • Permanent Blockchain Record
Scan with Saymatik Web3.0 Wallet

Saymatik Web3.0 Wallet