Skip to main content

> ML_DATASET // VISUAL-GENOME-DATASET_v1.0

Visual Genome: Connecting Language and Vision (Krishna et al. 2017)

Stanford University (Ranjay Krishna et al.) · Visual Relational Knowledge & Scene Graphs · 108,077 images with 5.4M region descriptions, 2.3M visual relationships, and 2.8M attributes

Visual Relational Knowledge & Scene GraphsCC-BY-4.0108,077 images with 5.4M region descriptions, 2.3M visual relationships, and 2.8M attributesopen

Dataset Profile & Characteristics

Label Type:Scene graphs: localized bounding boxes, attributes, and directed relational predicate triples
Languages:en
License Tier:permissive-open-source
Modalities:image, text, graph

Intended Use

  • Benchmarking scene graph generation, visual relationship detection, and grounded VQA

Prohibited / Discouraged Use

  • Facial recognition

Bias, Leakage & Privacy Risk Analysis

Privacy / Sensitive Data Risks:

Photographic images with human activities in public and private spaces.

Known Bias:

Heavy visual relationship long-tail (e.g., "on", "has", "wearing" dominate).

Known Benchmark Leakage:

Clean split mapped to MS COCO image identifiers.

Compatible Tools & Libraries