The Performance of ChatGPT-4o and DeepSeek-R1 in Interpreting Thyroid Nodule Ultrasound Text Reports: Multicenter Study
Background: Although thyroid nodules are detected in up to 60% of adults on ultrasound, the vast majority are benign, creating a substantial decision-making burden compounded by heterogeneous practice guidelines. Large language models (LLMs) show promise in processing unstructured medical text and are emerging as tools for report interpretation among both clinicians and patients. However, their reliability across distinct clinical tasks in thyroid ultrasound interpretation remains poorly charact
Record details
Published: 28 July 2026
Source: JMIR (Journal of Medical Internet Research)
Category: Research
Topics: Healthcare
Retrieved: 29 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Performance of 5 Large Language Models in Perioperative Consultation for Pediatric Hypospadias: Cross-Sectional Comparative Study
JMIR (Journal of Medical Internet Research) · 29 July 2026
Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety
arXiv cs.CY · 5 August 2026
PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong
arXiv · 10 April 2026
AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation
arXiv cs.HC · 28 July 2026
Institutional Trust and the Domestic AI Advantage: Evidence from DeepSeek and ChatGPT Users in China
arXiv · 31 May 2026
Can AI Make Conflicts Worse? An Alignment Failure in LLM Deployment Across Conflict Contexts
arXiv · 21 May 2026
How to cite this record
ethics.ai (28 July 2026), “The Performance of ChatGPT-4o and DeepSeek-R1 in Interpreting Thyroid Nodule Ultrasound Text Reports: Multicenter Study,” evidence record 14236, https://ethics.ai/record/14236 (originally published by JMIR (Journal of Medical Internet Research)).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.