Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Shahm_-_bart-german-8bits – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
Shahm_-_bart-german-8bits
like
0
transformers
safetensors
bart
text-generation
autotrain_compatible
endpoints_compatible
8-bit
bitsandbytes
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
bart-german - bnb 8bits
Model creator:
https://huggingface.co/Shahm/
Original model:
https://huggingface.co/Shahm/bart-german/
Original model description:
license: apache-2.0 tags:
generated_from_trainer
summarization datasets:
mlsum language: de metrics:
rouge model-index:
name: mode-bart-deutsch results:
task: name: Summarization type: summarization dataset: name: mlsum de type: mlsum args: de metrics:
name: Rouge1 type: rouge value: 41.698
mode-bart-deutsch
This model is a fine-tuned version of
facebook/bart-base
on the mlsum de dataset. It achieves the following results on the evaluation set:
Loss: 1.2152
Rouge1: 41.698
Rouge2: 31.3548
Rougel: 38.2817
Rougelsum: 39.6349
Gen Len: 63.1723
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
learning_rate: 5e-05
train_batch_size: 6
eval_batch_size: 6
seed: 42
optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
lr_scheduler_type: linear
num_epochs: 3.0
Training results
Framework versions
Transformers 4.16.0.dev0
Pytorch 1.10.0+cu111
Datasets 1.17.0
Tokenizers 0.10.3