A curated RL dataset of ~720 tasks with instructions, environments, and verifiers for agentic training.
OpenThoughts-Agent is an open-source effort to curate the best datasets for training agents. Our first release includes datasets, models and our research codebase;
OpenThinker-Agent-v1 is a model trained for agentic tasks such as Terminal-Bench 2.0 and SWE-Bench.
We built… See the full description on the dataset page:
https://huggingface.co/datasets/open-thoughts/OpenThoughts-Agent-v1-RL.