August 22, 2024Open Access

MedFrenchmark, a Small Set for Benchmarking Generative LLMs in Medical French

Puntos clave

Los puntos clave no están disponibles para este artículo en este momento.

Resumen

Generative Large Language Models (LLMs) have become ubiquitous in various fields, including healthcare and medicine. Consequently, there is growing interest in leveraging LLMs for medical applications, leading to the emergence of novel models daily. However, evaluation and benchmarking frameworks for LLMs are scarce, particularly those tailored for medical French. To address this gap, we introduce a minimal benchmark consisting of 114 open questions designed to assess the medical capabilities of LLMs in French. The proposed benchmark encompasses a wide range of medical domains, reflecting real-world clinical scenarios’ complexity. A preliminary validation involved testing seven widely used LLMs with a parameter size of 7 billion. Results revealed significant variability in performance, emphasizing the importance of rigorous evaluation before deploying LLMs in medical settings. In conclusion, we present a novel and valuable resource for rapidly evaluating LLMs in medical French. By promoting greater accountability and standardization, this benchmark has the potential to enhance trustworthiness and utility in harnessing LLMs for medical applications.

Connected Papers

Building similarity graph...

Analyzing shared references across papers

Discussion

Cite this study

Quercia et al. (Thu,) studied this question.

www.synapsesocial.com/papers/68e5b602b6db64358754f147 — DOI: https://doi.org/10.3233/shti240486

Authors

A. Quercia

Jamil Zaghir

Christian Lovis

Actions

Institutions

University of Geneva

References and Citations

Connected Papers

Building similarity graph...

Analyzing shared references across papers

MedFrenchmark, a Small Set for Benchmarking Generative LLMs in Medical French

Puntos clave

Resumen

Citation Network

Connected Papers

Discussion

Cite this study

Authors

Actions

Institutions

References and Citations

Citation Network

Connected Papers

Discussion

Also consider