Descripció
Dataset Card for Tilde-MODEL-Catalan
Dataset Summary
This dataset contains the German version of the Tilde-MODEL corpus aligned with a Catalan translation. The catalan text has been obtained using Apertium's RBMT system from the Spanish version. It contains 3.4M segments.
Supported Tasks and Leaderboards
This dataset can be used to train NMT and SMT systems. It has been used as a training corpus for the Softcatalà machine translation engine.
Languages… See the full description on the dataset page: https://huggingface.co/datasets/softcatala/Tilde-MODEL-Catalan.
Adreça de descàrrega:
https://huggingface.co/datasets/softcatala/Tilde-MODEL-Catalan| property | value | ||||||
|---|---|---|---|---|---|---|---|
| name | Catalan-German aligned corpora to train NMT systems. |
||||||
| description | |
||||||
| license |
|
||||||
| sameAs | https://www.softcatala.org/dades-obertes/tilde-model-catalan/ |
||||||
| url | https://huggingface.co/datasets/softcatala/Tilde-MODEL-Catalan |