This is a protein binder interaction dataset built from multiple design papers and supplementary information source files.
Each row is a target-binder pair with:
the target name and target sequence
the binder identifier and binder sequence
a binary interaction label
the paper/source of origin
source_publication: paper or source string associated with the interaction
target: target name
target_sequence: target amino-acid sequence
binder_id: binder… See the full description on the dataset page:
https://huggingface.co/datasets/yk0/litscrape.