Skip to content

Latest commit

 

History

History

euq-euq

opus-2020-10-04.zip

  • dataset: opus
  • model: transformer
  • source language(s): eng eus
  • target language(s): eng eus
  • model: transformer
  • pre-processing: normalization + SentencePiece (spm32k,spm32k)
  • a sentence initial language token is required in the form of >>id<< (id = valid target language ID)
  • download: opus-2020-10-04.zip
  • test set translations: opus-2020-10-04.test.txt
  • test set scores: opus-2020-10-04.eval.txt

Benchmarks

testset BLEU chr-F
Tatoeba-test.eng-eus.eng.eus 28.1 0.563
Tatoeba-test.eus-eng.eus.eng 42.6 0.606
Tatoeba-test.multi.multi 36.3 0.583

opus4m+btTCv20210807-2021-09-30.zip

Benchmarks

testset BLEU chr-F #sent #words BP
Tatoeba-test-v2021-08-07.multi-multi 40.6 0.615 2120 15237 0.959