The LDQuAd is a comprehensive dataset consisting of 154,000 question-answer pairs in Azerbaijani language. This dataset is the first of its kind in the Azerbaijani language at such a scale, and LocalDoc is proud to present it.
Dataset Features
Total Q&A Pairs: 154,000
Questions without Answers: Approximately 30% of the questions do not have answers. This design choice helps in training models to effectively… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/LDQuAd.