Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
EvoAgentBench provides standardized train/test splits across five diverse task domains for evaluating AI agent self-evolution. It was created by EverMind-AI and last updated on April 14, 2026. The benchmark is designed for reproducible comparison of skill extraction and experience reuse methods.
License is unknown; terms of use must be verified before application.