Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A benchmark dataset comprising over 14,500 questions on non-synthetic images, created to assess stereotype biases in Large Multimodal Models (LMMs). The dataset, authored by ucf-crcv, was last updated on May 16, 2025. It spans nine diverse domains and 54 sub-domains to rigorously evaluate LMM performance in visually grounded stereotypical scenarios.
License is unknown, which may restrict usage.