Re: [PATCH V2] scripts/spdxcheck.py: Strictly read license files in utf-8
Jonathan Corbet <[email protected]>
| Newsgroups | org.kernel.vger.linux-spdx,org.kernel.vger.linux-kernel |
|---|---|
| Message-ID | <[email protected]> |
Nishanth Menon <[email protected]> writes: > Commit bc41a7f36469 ("LICENSES: Add the CC-BY-4.0 license") > unfortunately introduced LICENSES/dual/CC-BY-4.0 in UTF-8 Unicode text > While python will barf at it with: > > FAIL: 'ascii' codec can't decode byte 0xe2 in position 2109: ordinal not in range(128) > Traceback (most recent call last): > File "scripts/spdxcheck.py", line 244, in <module> > spdx = read_spdxdata(repo) > File "scripts/spdxcheck.py", line 47, in read_spdxdata > for l in open(el.path).readlines(): > File "/usr/lib/python3.6/encodings/ascii.py", line 26, in decode > return codecs.ascii_decode(input, self.errors)[0] > UnicodeDecodeError: 'ascii' codec can't decode byte 0xe2 in position 2109: ordinal not in range(128) > > While it is indeed debatable if 'Licensor.' used in the license file > needs unicode quotes, instead, force spdxcheck to read utf-8. > > Reported-by: Rahul T R <[email protected]> > Signed-off-by: Nishanth Menon <[email protected]> > Reviewed-by: Thomas Gleixner <[email protected]> I've applied this, thanks. jon