MultiLS-Japanese is a lexical complexity prediction (LCP) and lexical simplification (LS) dataset for Japanese. The MultiLS-Japanese dataset was created by Adam Nohejl, Akio Haykawa, and Yusuke Ide.
A journal paper about the dataset.
More information and additional files on the MultiLS-Japanese Github repo.
Multilingual LS and LCP Data
Related datasets for 9 more languages: MLSP2024 dataset on Hugging Face Hub.