open-source-english-catalan-corpus

Compartiu

Descripció

Dataset Card for open-source-english-catalan-corpus

Dataset Summary

Translation memory built from more than 180 open source projects. These include LibreOffice, Mozilla, KDE, GNOME, GIMP, Inkscape and many others. It can be used as translation memory or as training corpus for neural translators.

Supported Tasks and Leaderboards

[More Information Needed]

Languages

Catalan (ca) English (en)

Dataset Structure

Data Instances

[More… See the full description on the dataset page: https://huggingface.co/datasets/softcatala/open-source-english-catalan-corpus.

Adreça de descàrrega:

https://huggingface.co/datasets/softcatala/open-source-english-catalan-corpus
Autor:

Softcatalà

Hugging Face:

39 baixades (darrers 30 dies)

1 «m'agrada»