Resolve license issue with the utf8proc code. #364

DennisHeimbigner · 2017-02-16T21:30:28Z

Re: Github issue #349.

Update utf8proc.[ch] to use the version now
maintained by the Julia Language project
(https://github.com/JuliaLang/utf8proc/blob/master/LICENSE.md).
The license for the previous version was
unacceptable for the Debian and Ubuntu release
systems. The new version both updates the code
and addresses the license issue.

It turns out that the utf8proc software we are using
was turned over to the Julia Language developers
and the license terms changed to allow modification.
(https://github.com/JuliaLang/utf8proc/blob/master/LICENSE.md).

So the fix here is as follows:

Wrap the library with a fixed interface: libdispatch/dutf8.c
and include/ncutf8.h.
Replace the existing utf8proc code with the new version
from https://github.com/JuliaLang/utf8proc.
Add a couple more test cases: nc_test/tst_utf8_validate.c
and nc_test_utf8_phrases.c. If/when I can find a usable
normalization test, I will incorporate that later.

Update utf8proc.[ch] to use the version now maintained by the Julia Language project (https://github.com/JuliaLang/utf8proc/blob/master/LICENSE.md). The license for the previous version was unacceptable for the Debian and Ubuntu release systems. The new version both updates the code and addresses the license issue. It turns out that the utf8proc software we are using was turned over to the Julia Language developers and the license terms changed to allow modification. (https://github.com/JuliaLang/utf8proc/blob/master/LICENSE.md). So the fix here is as follows: 1. Wrap the library with a fixed interface: libdispatch/dutf8.c and include/ncutf8.h. 2. Replace the existing utf8proc code with the new version from https://github.com/JuliaLang/utf8proc. 3. Add a couple more test cases: nc_test/tst_utf8_validate.c and nc_test_utf8_phrases.c. If/when I can find a usable normalization test, I will incorporate that later.

WardF · 2017-02-17T20:04:46Z

There are some build failures on Windows with Visual Studio. I'll get them resolved and then merge this pull request.

DennisHeimbigner · 2017-02-17T20:23:32Z

Sorry, I always forget to test with cmake. =Dennis

…

On 2/17/2017 1:04 PM, Ward Fisher wrote: There are some build failures on Windows with Visual Studio. I'll get them resolved and then merge this pull request. — You are receiving this because you authored the thread. Reply to this email directly, view it on GitHub <#364 (comment)>, or mute the thread <https://github.com/notifications/unsubscribe-auth/AA3P29mw4KuXxcNOpjkM3pLggH8eMnWHks5rdf1fgaJpZM4MDkUB>.

WardF · 2017-02-17T22:26:07Z

Not a problem. The root of the issue is described here, and should be easy enough for me to fix. Attending to it now.

https://msdn.microsoft.com/en-us/library/62688esh.aspx

…nidata/netcdf-c/pulls/364

DennisHeimbigner · 2017-02-18T19:11:31Z

For what its worth, I only use a couple of the functions in utf8proc. None of the others
need to be exported.

WardF · 2017-02-27T18:01:06Z

Propagating changes from master into branch, then working on resolving this.

This is a follow-on in that the old utf8 code was still being used in ncgen to convert utf8->utf16 when converting cdl to Java (see genj.c). The new code apparently has no utf16 support, but it does have utf32 support. Converting utf32 -> utf16 can be approximated by truncating the 32bits to 16 bits, unless the top 16 bits are not zero. This latter condition is unlikely to be common because it implies use of some rather obscure characters. So solution is to convert to utf32 and truncate to 16 bits to get utf16. An error is reported if the high-order truncated 16 bits are not zero. If we get complaints, then I will figure out how to convert full utf32 to a utf16 pair. Also removed the old code from ncgen.

This is a follow-on in that the old utf8 code was still being used in ncgen to convert utf8->utf16 when converting cdl to Java (see genj.c). The new code apparently has no utf16 support, but it does have utf32 support. Converting utf32 -> utf16 can be approximated by truncating the 32bits to 16 bits, unless the top 16 bits are not zero. This latter condition is unlikely to be common because it implies use of some rather obscure characters. So solution is to convert to utf32 and truncate to 16 bits to get utf16. An error is reported if the high-order truncated 16 bits are not zero. If we get complaints, then I will figure out how to convert full utf32 to a utf16 pair. Other changes: 1. removed the old code from ncgen. 2. changed UTF8PROC_DLLEXPORT (in utf8proc) to EXTERNL and added appropriate includes. This should fix issue #404, but since we cannot duplicate the failure, I am not quite sure.

DennisHeimbigner and others added 3 commits February 16, 2017 14:27

Update RELEASENOTES

de0b55d

Updated CMakeLists.txt to work on Windows.

7cce9c3

Updated for Visual Studio support, in support of https://github.com/U…

da28564

…nidata/netcdf-c/pulls/364

Merging master into branch.

78c0f34

WardF merged commit 1603cd0 into master Feb 27, 2017

WardF deleted the utffix.dmh branch February 27, 2017 19:47

WardF mentioned this pull request Jun 6, 2017

License problem: ConvertUTF is non-free, use libicu instead #349

Closed

DennisHeimbigner mentioned this pull request Jun 19, 2017

Remove more old utf8 code #430

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Resolve license issue with the utf8proc code. #364

Resolve license issue with the utf8proc code. #364

DennisHeimbigner commented Feb 16, 2017

WardF commented Feb 17, 2017

DennisHeimbigner commented Feb 17, 2017 via email

WardF commented Feb 17, 2017

DennisHeimbigner commented Feb 18, 2017

WardF commented Feb 27, 2017

Resolve license issue with the utf8proc code. #364

Resolve license issue with the utf8proc code. #364

Conversation

DennisHeimbigner commented Feb 16, 2017

WardF commented Feb 17, 2017

DennisHeimbigner commented Feb 17, 2017 via email

WardF commented Feb 17, 2017

DennisHeimbigner commented Feb 18, 2017

WardF commented Feb 27, 2017