This dataset is based on Fleurs from Google. We matched the English sentences with Norwegian sentences and formatted it to an Alpaca-style dataset.
{
"instruction": "Oversett teksten fra engelsk til norsk",
"input": "English string",
"output": "Norwegian string"
}
This dataset was created by Ruter during Ruter's AI Lab effort to fine-tune LLaMA-2 models for Norwegian.
Following the original dataset… See the full description on the dataset page:
https://huggingface.co/datasets/RuterNorway/Fleurs-Alpaca-EN-NO.