Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Persian-language evaluation results for the openai/gpt-oss-120b model on a suite of benchmarks. The dataset covers general knowledge, mathematical reasoning, and language understanding tasks, using the first 15 samples per sub-task with 0-shot prompting. It was created by author artindnr and last updated on Hugging Face in July 2026.
License is unknown; users should verify terms before use.