Loading...
Loading...
AgentDoG-Lite provides a test set for evaluating AI agent safety based on multi-turn execution trajectories. The dataset, created by AI45Research and last updated in July 2026, is designed for a specific challenge assessing whether an agent performs unsafe actions or follows unsafe decision patterns. It focuses on a binary safety judgment task for complete agent interaction sequences.
License is unknown; users should verify terms before use.