3 ms·
Unicode is a terrible mess. I remember reading about it with enthusiasm in the nineties, but it didn't take long to realize the committee of designers simply p
by zerofan 10y ago
Unicode is a terrible mess. I remember reading about it with enthusiasm in the nineties, but it didn't take long to realize the committee of designers simply punted on every difficult decision instead of making a stand. They dumped everything on the programmers. At this point, I think it would be less work to convert 7 billion people to using ascii than it would to enable 10 million programmers to use that pile of a standard robustly. History lessons won't help us here.
- js8 10y ago> committee of designers simply punted on every difficult decision instead of making a stand Are you sure that it was a bad thing? If they made a stand, we would now had to use (arguably worse) UCS2 or UCS4 encoding instead of UTF-8 (which de facto won).
- zerofan 10y agoUTF-8 is a nifty compression scheme - like a Huffman encoding that doesn't require a table to implement, but the committee doesn't deserve any credit for that (unless Thompson and Pike were on the committee, which I doubt). Besides, the reason most of us actually like UTF-8 is because it leaves ascii alone (which is all I ever use) while pretending to handle the general case. It doesn't help end users or programmers deal with any of the nonsense around multiple ways to encode glyphs (combining codes vs accented codes), deal with surrogates (yes, people encode surrogates in UTF-8), lexical sorting, or anything else. I'll bet there are dozens of incompatible ways strings are UTF-8 encoded in the real world, each of them a bug for interoperability, and all of that blame falls on Unicode being a terrible standard. So yes, I'm sure.
- deleted 10y ago[deleted]