SAQR is a benchmark dataset for Arabic Handwritten Text Recognition (HTR), Printed-to-Handwritten Retrieval/Matching, and Demographic (Gender) Analysis.
The dataset contains 2,263 aligned pairs of synthetic printed ground-truth crops and their handwritten counterpart line images written by diverse writers across 3 official splits (Train: 1,575, Validation: 335, Test: 353).