Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
1.9 million data annotations in JSON format, created by OpenGVLab for instruction-tuning multimodal AI models. The dataset was last updated in June 2024. It is derived from image datasets like M3IT, which underwent quality filtering processes.
Images and videos are not included in the provided JSON files; users must follow separate instructions to obtain the corresponding media. The license is tagged as 'mit' but the specific terms are not detailed in the input.