PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 25, 2024European Archives of Oto-Rhino-Laryngology37 citationsOpen Access

Reliability of large language models for advanced head and neck malignancies management: a comparison between ChatGPT 4 and Gemini Advanced

View Full Paper
ALAndrea De LorenziGPGiorgia PuglieseAMAntonino Maniaci

Key Points

Key points are not available for this paper at this time.

Abstract

Abstract Purpose This study evaluates the efficacy of two advanced Large Language Models (LLMs), OpenAI’s ChatGPT 4 and Google’s Gemini Advanced, in providing treatment recommendations for head and neck oncology cases. The aim is to assess their utility in supporting multidisciplinary oncological evaluations and decision-making processes. Methods This comparative analysis examined the responses of ChatGPT 4 and Gemini Advanced to five hypothetical cases of head and neck cancer, each representing a different anatomical subsite. The responses were evaluated against the latest National Comprehensive Cancer Network (NCCN) guidelines by two blinded panels using the total disagreement score (TDS) and the artificial intelligence performance instrument (AIPI). Statistical assessments were performed using the Wilcoxon signed-rank test and the Friedman test. Results Both LLMs produced relevant treatment recommendations with ChatGPT 4 generally outperforming Gemini Advanced regarding adherence to guidelines and comprehensive treatment planning. ChatGPT 4 showed higher AIPI scores (median 3 2–4) compared to Gemini Advanced (median 2 2–3), indicating better overall performance. Notably, inconsistencies were observed in the management of induction chemotherapy and surgical decisions, such as neck dissection. Conclusions While both LLMs demonstrated the potential to aid in the multidisciplinary management of head and neck oncology, discrepancies in certain critical areas highlight the need for further refinement. The study supports the growing role of AI in enhancing clinical decision-making but also emphasizes the necessity for continuous updates and validation against current clinical standards to integrate AI into healthcare practices fully.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lorenzi et al. (2024) studied this question.

synapsesocial.com/papers/68e686b9b6db64358760f0fbhttps://doi.org/10.1007/s00405-024-08746-2
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Performance and Consistency of ChatGPT‐4 Versus Otolaryngologists: A Clinical Case Series2024 · 42 citations
  2. 2Validation of the Quality Analysis of Medical Artificial Intelligence (QAMAI) tool: a new tool to assess the quality of health information provided by AI platforms2024 · 88 citations
  3. 3Generative artificial intelligence in otolaryngology–head and neck surgery editorial: be an actor of the future or follower2024 · 7 citations
  4. 4Artificial Intelligence in Head and Neck Cancer: A Systematic Review of Systematic Reviews2023 · 74 citations
  5. 5Accuracy of ChatGPT‐Generated Information on Head and Neck and Oromaxillofacial Surgery: A Multicenter Collaborative Analysis2023 · 122 citations