This dataset consists of templated sentences with the masked word being sensitive to race, e.g., African.
See GENTER and GRADIEND Religion Data for similar datasets.
The dataset uses one subset per class. Subset names are class identifiers: white, black, asian. Each subset has columns masked, split, and one column per class (e.g. white, black, asian) giving the token for that class in that row.
from datasets import load_dataset