PLLuMIC - Polish Large Language Model (PLLuM) Instruction Corpus
Dataset Details
Dataset Description
We release the first representative subset of the PLLuM Instruction Corpus (PLLuMIC), which we believe to be useful in guiding and planning the development of similar LLM datasets. PLLuMIC is a hand-crafted set of LLM fine-tuning Polish language instructions, developed in line with the annotation guidelines and covering a functional… See the full description on the dataset page: https://huggingface.co/datasets/pelcra/PLLuMIC.