Dataset Card for HistCiph — Polish
Dataset Description
Dataset Summary
The Polish subset of HistCiph is part of the first publicly available multilingual collection of historically grounded plaintext–ciphertext pairs for classical homophonic substitution ciphers. It pairs diachronically balanced historical Polish plaintext with independently generated homophonic substitution keys and controlled transcription noise, producing four distinct ciphertext variants… See the full description on the dataset page: https://huggingface.co/datasets/mbruton/polish_encrypted_HistCiph.