Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
TRM-Preference is a dataset introduced in the paper 'Characterizing, Evaluating, and Optimizing Complex Reasoning' for evaluating and optimizing reasoning traces in Large Reasoning Models. The dataset applies the ME² principle to assess 'how a model thinks' across dimensions like Macro-Efficiency. It was authored by zzzhr97 and last updated on Hugging Face in June 2026.
License is unknown; terms of use must be verified before application.