Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Over 64,000 curated temporal segments from unconstrained, real-world YouTube videos. WildVid-LIP is a large-scale, open-source dataset providing precise timestamp anchors for training Visual Speech Recognition and multimodal models. The dataset was created by Rizul2159 and was last updated on June 16, 2026.
License is unknown; terms of use must be verified before application.