Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
AMALIA-VL-DPO Dataset is a collection of Direct Preference Optimization (DPO) data for the AMALIA-VL project. It contains preference triplets where each row includes a prompt with normalized role-content turns and image placeholders, a chosen assistant response, and a rejected response. The dataset was created by author 'amalia-llm' and was last updated on June 30, 2026.
License is unknown, which may restrict usage. The full description is hosted externally on the Hugging Face dataset page.