visualgenome¶
| Field | Value |
|---|---|
| Description | Visual Genome is a dataset, a knowledge base, an ongoing effort to connect structured image concepts to language. |
| Folder | /datasets/ai/visualgenome |
| Discipline | AI / computer vision / multimodel AI |
| DOI | 10.1007/s11263-016-0981-7 |
| Link | Access Data |
| Public | True |
| Publication Date | 2017-02-06 |
| Downloaded | 2026-01-07 |
| Data Type | LMDB, SquashFS, Extracted JPG files on Ceph |
| Dataset Size | 21G (extracted) |
| Number of Files | 108,265 (extracted) |
| Usage | $ module avail |
| Usage Policy Link | http://creativecommons.org/licenses/by/4.0 |
| Usage Policy | This dataset is distributed under the Creative Commons Attribution 4.0 International (CC BY 4.0) license. Users may use, share, adapt, and redistribute the data with appropriate credit to the dataset authors and hosting repository. A link to the license must be included (https://creativecommons.org/licenses/by/4.0/), and any modifications should be clearly indicated. Users are responsible for ensuring that the dataset is appropriate for their applications. |
| Citation | Krishna, R., Zhu, Y., Groth, O. et al. Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations. Int J Comput Vis 123, 32–73 (2017). https://doi.org/10.1007/s11263-016-0981-7 |
| BibTeX | 📜 View BibTeX citation@article{Krishna2017, |
Ceph access¶
This dataset is also available in raw/extracted form on RCAC Ceph/S3-compatible object storage.
| Parameter | Value |
|---|---|
| Endpoint | https://s3.anvil.rcac.purdue.edu |
| Bucket | ai-datasets |
| Access | Public read-only |
For detailed instructions, see the AI datasets overview.