A collection of 219 heterodimers from dockground benchmark 4 dataset, and 1503 heterodimeric structures from a recent study (Green, A. G. et al. Nat. Commun. 12, 1–12 (2021)) (dubbed “marks”). The benchmark dataset was used to train an original model, that model was used to predict the structures in the second dataset.
This dataset is split into three subsets, each with two splits corresponding to the groups used in the source (marks, dockground).
Quickstart… See the full description on the dataset page: https://huggingface.co/datasets/RosettaCommons/FoldDock.