4 ms·
> for this reason, String iteration should be based on codepoints This leads you to the problem where you'll get different results iterating over n a ï v e v
by nerdbert 3y ago
> for this reason, String iteration should be based on codepoints
This leads you to the problem where you'll get different results iterating over
n a ï v e
vs
n a ̈ i v e
And I can't see how that's ever going to be a useful outcome.
If you normalize everything first, then you can sidestep this to some degree, but then in effect your normalization has turned codepoint iteration into grapheme iteration for most common Latin-script text characters.