This dataset consists of court cases from the Republic of Azerbaijan.
Overview
It was formed based on 1,200,000 court cases.
The data has been preliminarily normalized and split into sentences.
The dataset consists of 37 million sentences and approximately 500-600 million tokens.
Dataset Structure
Each row represents a single sentence extracted from a court case document.