Beyond English: ChatGPT's instructions across EU languages
Identifikátory výsledku
Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F60460709%3A41110%2F24%3A100014" target="_blank" >RIV/60460709:41110/24:100014 - isvavai.cz</a>
Výsledek na webu
<a href="https://doi.org/10.1016/j.resuscitation.2024.110367" target="_blank" >https://doi.org/10.1016/j.resuscitation.2024.110367</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.1016/j.resuscitation.2024.110367" target="_blank" >10.1016/j.resuscitation.2024.110367</a>
Alternativní jazyky
Jazyk výsledku
angličtina
Název v původním jazyce
Beyond English: ChatGPT's instructions across EU languages
Popis výsledku v původním jazyce
This study evaluates ChatGPT's ability to provide cardiopulmonary resuscitation (CPR) instructions across 24 official European Union languages. Using GPT-4o, three responses were generated per language and evaluated by three large language models on medical accuracy, comprehensiveness, clarity, safety, and emergency suitability. While English, Dutch, and Polish demonstrated optimal performance, and some languages like Greek showed more suboptimal ratings, no instructions were rated as critically unsafe. The findings suggest consistent baseline reliability across languages while highlighting the need for comprehensive multilingual benchmarks in healthcare AI systems.
Název v anglickém jazyce
Beyond English: ChatGPT's instructions across EU languages
Popis výsledku anglicky
This study evaluates ChatGPT's ability to provide cardiopulmonary resuscitation (CPR) instructions across 24 official European Union languages. Using GPT-4o, three responses were generated per language and evaluated by three large language models on medical accuracy, comprehensiveness, clarity, safety, and emergency suitability. While English, Dutch, and Polish demonstrated optimal performance, and some languages like Greek showed more suboptimal ratings, no instructions were rated as critically unsafe. The findings suggest consistent baseline reliability across languages while highlighting the need for comprehensive multilingual benchmarks in healthcare AI systems.
Klasifikace
Druh
J<sub>imp</sub> - Článek v periodiku v databázi Web of Science
CEP obor
—
OECD FORD obor
30221 - Critical care medicine and Emergency medicine
Návaznosti výsledku
Projekt
—
Návaznosti
S - Specificky vyzkum na vysokych skolach
Ostatní
Rok uplatnění
2024
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů
Údaje specifické pro druh výsledku
Název periodika
RESUSCITATION
ISSN
0300-9572
e-ISSN
0300-9572
Svazek periodika
202
Číslo periodika v rámci svazku
110367
Stát vydavatele periodika
CZ - Česká republika
Počet stran výsledku
2
Strana od-do
—
Kód UT WoS článku
001310334700001
EID výsledku v databázi Scopus
2-s2.0-85202652226