This repository provides four multilingual, augmented, and similarity-enhanced variants of the Conceptual Captions 3M (CC3M) dataset.The goal is to support research in vision–language modeling, multimodal alignment, data augmentation, and low-resource language evaluation.