Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
A 3-hour collection of naturalistic Chinese-English code-mixed speech sourced from social media videos. The dataset includes dual-level annotations, featuring manual token-level labeling for prosodic analysis at switch boundaries. It was created by hafsamenaz1 and last updated on Hugging Face in April 2026.
License is unknown; terms of use must be verified.