This dataset contains a collection of three text subsets designed for instruction tuning and evaluation of large language models (LLMs). The subsets provide examples across Japanese language instruction and mathematical reasoning tasks.
Dataset Details
Dataset Description
This dataset consists of three subsets:
Ichikara
Focus: Japanese language instruction for LLMs.
Provenance: Created by researchers at RIKEN and collaborators for supporting… See the full description on the dataset page: https://huggingface.co/datasets/arcee-ai/DAM.