Re: Machine translation. Was: Re: Translation platform migration: status update
Piotr Drąg <[email protected]>
| Newsgroups | gmane.linux.redhat.internationalization |
|---|---|
| Message-ID | <CAE4jGzrw23G6ybCgiviB-2+92qTkFpw0gJsVgkibWhGDHDr4Ow@mail.gmail.com> |
niedz., 18 sie 2019 o 20:48 Jean-Baptiste <[email protected]> napisał(a): > > As all we do is open source, it's fine to reuse our work for any activity as long as it respect the license. > > If such activities interests you, here is an extraction of all our translation I did for the amagama team: > > https://jibecfed.fedorapeople.org/partage/2019-07-06-all-fedora-zanata-org-tm.7z > > And here is the script to run it yourself: > https://github.com/Jibec/fedora-translation-statistics/blob/master/get_zanata_tm.py > > You just have to change the headers; > headers = {"X-Auth-User":"jibecfed", "X-Auth-Token":"change-me"} > > Note: it produce one multilingual translation memory file (tmx) per project-iteration. If you had a Python script to transform+merge it in one tmx per language with all projects, it would be great! > > Would there be community objections to use of the translations as > machine translation training material? If not, this might allow for > some collaboration with machine translation companies in a similar way > to the use of reCaptcha for crowdsourcing image labels. I’m pretty sure translators have copyright over their work and you can’t reuse translations without attribution (additionally to any license concerns.) I’m not a lawyer, of course. Best regards, -- Piotr Drąg https://piotrdrag.fedorapeople.org _______________________________________________ trans mailing list -- [email protected] To unsubscribe send an email to [email protected] Fedora Code of Conduct: https://docs.fedoraproject.org/en-US/project/code-of-conduct/ List Guidelines: https://fedoraproject.org/wiki/Mailing_list_guidelines List Archives: https://lists.fedoraproject.org/archives/list/[email protected]