Dataset Card for Human PeptideAtlas 2025-01 Peptides
Dataset Summary
Random split of the distinct peptide sequences released in the Peptide sequences in FASTA format file (125 MB) from the Human PeptideAtlas 2025-01 build.Each record is just the amino-acid sequence (uppercase 20-AA alphabet + “U”, “O”). No headers, spectra, or metadata are included.
Reference
Desiere et al., "The PeptideAtlas Project", Nucleic Acids Research, 2006, 34, D655-D658