Data as of Jul 25, 2026 · Based on 2,718,867 AI responses across 9,511 prompts · See how Parse measures this
Visual Commonsense Reasoning (VCR) is a task and large-scale dataset for cognition-level visual understanding, requiring models to answer challenging visual questions and provide rationales explaining their answers. The dataset includes 290k multiple-choice questions with correct answers and rationales, built on 110k images and 80 object categories from COCO.
Parse Score