Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
October 15, 2025Open Access

Performance Comparison of Human Doctors and Large Language Models in Tuberculosis Triage, Diagnosis, and Management:An Experimental Study (Preprint)

View Full Paper
Ask AI
Bookmark
Share

Authors

JLJin LiaoWHWenjun HeHPH.L. Pan

Discussion

Loading...

Member takes

Overview

Experimental study evaluates large language models against TB physicians for diagnosis and management, indicating potential efficiency improvements.

Key Points

  • Large language models achieved similar precision and higher recall than human physicians in tuberculosis management tasks.
  • The study assessed 17 standardized tuberculosis cases using precision, recall, and F1 scores to evaluate performance.
  • LLMs outperformed physicians in subjective evaluations of suitability and information quality, though physician responses were more readable.
  • Integrating LLMs may enhance decision efficiency in tuberculosis care without replacing human doctors.

Cite This Study

Liao et al. (2025) studied this question.

synapsesocial.com/papers/68efd921056559ef428774c2https://doi.org/10.2196/preprints.85613
View Full Paper
Ask AI
Bookmark
Share

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Large Language Model Symptom Identification From Clinical Text: Multicenter Study2025 · 14 citations
  2. 2Application of Large Language Models in Complex Clinical Cases: Cross-Sectional Evaluation Study2025 · 5 citations
  3. 3Using Large Language Models to Analyze Symptom Discussions and Recommendations in Clinical Encounters .2025
  4. 4Large language models for automatable real-world performance monitoring of diagnostic decision support systems: a comparison to manual doctor panel review in a prospective clinical study2025
  5. 5Improving tuberculosis-related knowledge in tuberculosis patients: Protocol for the development and validation of an evidence-based Q&A robot powered by large language models2025