Views
No views yet
Helsinki-NLP/opus-mt-en-es, packaged for offline use in Playto.Helsinki-NLP/opus-mt-en-es — MarianMT, transformer architecture1pip install ctranslate2 transformers sentencepiece
2ct2-transformers-converter \
3 --model Helsinki-NLP/opus-mt-en-es \
4 --output_dir opus-mt-en-es \
5 --quantization int8 \
6 --copy_files source.spm target.spm
7tar czf opus-mt-en-es-ct2.tar.gz opus-mt-en-es-ct2opus-mt-en-es-ct2.tar.gz)| File | Approx Size | Purpose |
|---|---|---|
model.bin | ~61MB | CTranslate2 int8 quantized weights |
shared_vocabulary.json | ~1.5MB | CTranslate2 vocab |
source.spm | ~800 KB | SentencePiece source tokenizer |
target.spm | ~800 KB | SentencePiece target tokenizer |
config.json | ~250 B | CTranslate2 config |
ctranslate2 (Python)1import ctranslate2
2import sentencepiece
3
4translator = ctranslate2.Translator("opus-mt-en-es", device="cpu", compute_type="int8")
5sp_source = sentencepiece.SentencePieceProcessor("opus-mt-en-es/source.spm")
6sp_target = sentencepiece.SentencePieceProcessor("opus-mt-en-es/target.spm")
7
8source_tokens = sp_source.encode("Hello, how are you?", out_type=str) + ["</s>"]
9results = translator.translate_batch([source_tokens])
10print(sp_target.decode(results[0].hypotheses[0]))
11# → "Hola, ¿cómo estás?"ct2rs (Rust)1use ct2rs::{Translator, Tokenizer};
2
3let tokenizer = Tokenizer::new("opus-mt-en-es")?;
4let translator = Translator::with_tokenizer("opus-mt-en-es", tokenizer, /* config */)?;
5let result = translator.translate_batch(&["Hello, how are you?".to_string()], /* options */)?;</s> appended to source token sequences. The ct2rs::Tokenizer wrapper handles this automatically; raw SentencePiece calls must add it manually.Helsinki-NLP/opus-mt-en-es. License is CC-BY 4.0 inherited from upstream.Helsinki-NLP. OPUS-MT — Open Machine Translation Models.
https://github.com/Helsinki-NLP/Opus-MT