Machine translation. Was: Re: Translation platform migration: status update
Jean-Baptiste <[email protected]>
| Newsgroups | gmane.linux.redhat.internationalization |
|---|---|
| Message-ID | <[email protected]> |
As all we do is open source, it's fine to reuse our work for any activity as long as it respect the license.
If such activities interests you, here is an extraction of all our translation I did for the amagama team:
https://jibecfed.fedorapeople.org/partage/2019-07-06-all-fedora-zanata-org-tm.7z
And here is the script to run it yourself:
https://github.com/Jibec/fedora-translation-statistics/blob/master/get_zanata_tm.py
You just have to change the headers;
headers = {"X-Auth-User":"jibecfed", "X-Auth-Token":"change-me"}
Note: it produce one multilingual translation memory file (tmx) per project-iteration. If you had a Python script to transform+merge it in one tmx per language with all projects, it would be great!
Would there be community objections to use of the translations as
machine translation training material? If not, this might allow for
some collaboration with machine translation companies in a similar way
to the use of reCaptcha for crowdsourcing image labels.
_______________________________________________
trans mailing list -- [email protected]
To unsubscribe send an email to [email protected]
Fedora Code of Conduct: https://docs.fedoraproject.org/en-US/project/code-of-conduct/
List Guidelines: https://fedoraproject.org/wiki/Mailing_list_guidelines
List Archives: https://lists.fedoraproject.org/archives/list/[email protected]