This Dataset is curated from the original Aya-Collection dataset that was open-sourced by Cohere under the Apache-2.0 license.
The Aya Collection is a massive multilingual collection comprising 513 million instances of prompts and completions that cover a wide range of tasks. This collection uses instruction-style templates from fluent speakers and applies them to a curated list of datasets. It also includes translations of instruction-style datasets into 101 languages.… See the full description on the dataset page:
https://huggingface.co/datasets/Cognitive-Lab/Aya_Gujarati.