Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
33,116 short, atomic manipulation episodes collected for pretraining vision-language-action models. The dataset contains 1.52 million frames and 9,532 language tasks, formatted for the LeRobot v2.1 framework and derived from Franka robot demonstrations. It was created by MINT-SJTU and last updated on July 8, 2026.
License is unknown; users should verify terms before use.