This model was trained on
Fineweb2 Ar sample dataset.
The tokenizer was also trained using the same dataset.
See
sample code
(usage and training) and
initial post
Updated: Jan. 12, 2025.
ModernBERT Arabic (MLM) experiment.
Educational and explorational uses only. Limited data, not fully trained.
Evaluation on 5% of the data, uses 2 GPUs.