Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
VLM-150M is a large-scale image-text dataset recaptioned using an SFT-enhanced Qwen2VL model to improve the alignment and detail of textual descriptions. The dataset was created by zhixiangwei and was last updated on July 28, 2025. Its repository is hosted at https://zxwei.site/hqclip/.
License is unknown; users should verify usage rights.