Logo del repository
  1. Home
 
Opzioni

Reliability of large language models for advanced head and neck malignancies management: a comparison between ChatGPT 4 and Gemini Advanced

Lorenzi, Andrea
•
Pugliese, Giorgia
•
Maniaci, Antonino
altro
Saibene, Alberto Maria
2024
  • journal article

Periodico
EUROPEAN ARCHIVES OF OTO-RHINO-LARYNGOLOGY AND HEAD & NECK
Abstract
Purpose: This study evaluates the efficacy of two advanced Large Language Models (LLMs), OpenAI's ChatGPT 4 and Google's Gemini Advanced, in providing treatment recommendations for head and neck oncology cases. The aim is to assess their utility in supporting multidisciplinary oncological evaluations and decision-making processes. Methods: This comparative analysis examined the responses of ChatGPT 4 and Gemini Advanced to five hypothetical cases of head and neck cancer, each representing a different anatomical subsite. The responses were evaluated against the latest National Comprehensive Cancer Network (NCCN) guidelines by two blinded panels using the total disagreement score (TDS) and the artificial intelligence performance instrument (AIPI). Statistical assessments were performed using the Wilcoxon signed-rank test and the Friedman test. Results: Both LLMs produced relevant treatment recommendations with ChatGPT 4 generally outperforming Gemini Advanced regarding adherence to guidelines and comprehensive treatment planning. ChatGPT 4 showed higher AIPI scores (median 3 [2-4]) compared to Gemini Advanced (median 2 [2-3]), indicating better overall performance. Notably, inconsistencies were observed in the management of induction chemotherapy and surgical decisions, such as neck dissection. Conclusions: While both LLMs demonstrated the potential to aid in the multidisciplinary management of head and neck oncology, discrepancies in certain critical areas highlight the need for further refinement. The study supports the growing role of AI in enhancing clinical decision-making but also emphasizes the necessity for continuous updates and validation against current clinical standards to integrate AI into healthcare practices fully.
DOI
10.1007/s00405-024-08746-2
WOS
WOS:001234603400003
Archivio
https://hdl.handle.net/11368/3082160
info:eu-repo/semantics/altIdentifier/scopus/2-s2.0-85194351387
https://link.springer.com/article/10.1007/s00405-024-08746-2
https://www.ncbi.nlm.nih.gov/pmc/articles/PMC11392976/
Diritti
open access
license:creative commons
license uri:http://creativecommons.org/licenses/by/4.0/
FVG url
https://arts.units.it/bitstream/11368/3082160/3/s00405-024-08746-2.pdf
Soggetti
  • Artificial intelligen...

  • Computer-assisted dia...

  • Head and neck cancer

  • Head and neck oncolog...

  • Large language model

  • Laryngeal carcinoma

  • Nasopharyngeal carcin...

  • Oncological diagnosi

  • Oropharyngeal carcino...

  • Parotid carcinoma

  • Tongue carcinoma

google-scholar
Get Involved!
  • Source Code
  • Documentation
  • Slack Channel
Make it your own

DSpace-CRIS can be extensively configured to meet your needs. Decide which information need to be collected and available with fine-grained security. Start updating the theme to match your nstitution's web identity.

Need professional help?

The original creators of DSpace-CRIS at 4Science can take your project to the next level, get in touch!

Realizzato con Software DSpace-CRIS - Estensione mantenuta e ottimizzata da 4Science

  • Impostazioni dei cookie
  • Informativa sulla privacy
  • Accordo con l'utente finale
  • Invia il tuo Feedback