EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation
While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when faced with non-English languages and multimodal diagnostic tasks. This study protocol describes the development of EuropeMedQA, the first comprehensive, multilingual, and multimodal medical examination dataset sourced from official regulatory exams in Italy, France, Spain, and Portugal. Following FAIR data principles and SPIRIT-AI guidelines, we des
Record details
Published: 15 April 2026
Source: arXiv
Category: Research
Topics: Regulation · Healthcare
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
First, Do No Harm (With LLMs): Mitigating Racial Bias via Agentic Workflows
arXiv · 20 April 2026
AI Agents Under EU Law
arXiv · 6 April 2026
AEGIS: An Operational Infrastructure for Post-Market Governance of Adaptive Medical AI Under US and EU Regulations
arXiv · 20 March 2026
Ethics and EU AI Act in Cases of Work Disability Risk and Alzheimer's Disease Risk Prediction
arXiv · 7 June 2026
The European Health Data Space: an opportunity to strengthen citizen rights and engage citizens in health data governance
OpenAlex · 20 January 2026
Shadow AI in Swedish Health Care: Qualitative Analysis of Physicians’ Free-Text Answers
JMIR (Journal of Medical Internet Research) · 28 July 2026
How to cite this record
ethics.ai (15 April 2026), “EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation,” evidence record 5813, https://ethics.ai/record/5813 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.