AraImgStory-10K is a custom dataset prepared for a Master's thesis project on retrieval-augmented multimodal story generation for Arabic image-based storytelling.
The dataset contains approximately 10,000 real images with extracted visual elements, emotion labels, generated Arabic stories, cleaned story texts, train/validation/test splits, a trained retriever model, a FAISS story index, and evaluation outputs.
The dataset supports research on Arabic image-based storytelling, multimodal learning, retrieval-augmented generation, Arabic natural language generation, and low-resource Arabic vision-language applications.
The archive includes Python scripts, processed dataset files, the trained retrieval model, FAISS index files, story metadata, and evaluation results.