This repository contains a sample of a larger dataset consisting of 15,200 structured and cleaned Arabic–German legal sentences distributed across 5 Excel files.
Full Dataset
The full dataset includes:
15,200 manually curated sentences
5 Excel files (XLSX format)
UTF‑8 encoding
No duplicates
Clean and ready for NLP, AI training, and text analysis
This repository contains a sample only.
Purchase the Full… See the full description on the dataset page: https://huggingface.co/datasets/adam672/arabic-german-legal-dataset.