We introduce a fully open-source suite designed for effective offline deep research agent training. DeepForge series includes collection of 66k QA pairs, 33k SFT trajectories, and 21k DPO pairs.
If you use DeepForge dataset in your research, please cite:
@article{zhou2026offseeker,
title={OffSeeker: Online Reinforcement Learning Is Not All You Need for Deep Research Agents},
author={Zhou, Yuhang and Zheng, Kai and Chen, Qiguang and Hu, Mengkang and… See the full description on the dataset page:
https://huggingface.co/datasets/OffSeeker/DeepForge.