This dataset is a format conversion from its original v1.0.0 format and released here under the same CC-BY-SA 4.0 license and conditions.
It contains Japanese instruction-like data intended for LLM construction/tuning.
The dataset only contains a 'train' split, with ~2.46M rows of data.
Thanks Jian Wu (@wujian123) for the help in converting and validating the dataset.