This is a conversion of BibleNLP corpus to the Parquet format,
The dataset contains partial and complete Bible translations in 835 languages, aligned by verse. Each language is stored as a separate Parquet file (eng.parquet, fra.parquet, …).
This format is derived from the eBible corpus corpus.json and is intended for fast columnar loading with Hugging Face datasets, Polars, Pandas, or DuckDB.
835 ISO 639-3… See the full description on the dataset page:
https://huggingface.co/datasets/nordpolemil/biblenlp-corpus.