This dataset is a multilingual and multimodal medical VQA benchmark extended from the WorldMedQA-V dataset. It evaluates the medical reasoning performance and cross-lingual consistency of VLMs across five languages: English, Korean, Japanese, Arabic, and Wolof.
Images: Provided in .parquet format for efficient loading and high accessibility.