Paper | GitHub
Large-scale, non-redundant benchmark for protein fold classification built from
Encyclopedia of Domains (TED) annotations
projected onto the Foldseek-clustered AlphaFold Database.
This dataset was presented in the paper Protein Fold Classification at Scale: Benchmarking and Pretraining.