7 ms·
> All screen readers should be able to read all Unicode symbols properly. Really? https://en.wikipedia.org/wiki/Cuneiform_Numbers_and_Punctuation https://en.w
by usrbinbash 4y ago
> All screen readers should be able to read all Unicode symbols properly.
Really?
https://en.wikipedia.org/wiki/Cuneiform_Numbers_and_Punctuation https://en.wikipedia.org/wiki/Cuneiform_Numbers_and_Punctuat...
How should 𒐉 be read exactly? Should it just say "Cuneiform Numeric Sign Four Dish"? Should it say 4? Should it say 4, but in Arameic? How should it deal with the fact that the system didn't have a symbol for the Radix Point, and so the actual value of numbers depends on context?
- account42 4y agoI don't think we need to worry about cuneiform yet when we have the current lack of even basic support described in the article. Reading the unicode character name out would be a good fallback for less important cases - certainly be better than skipping over characters. Let's not make perfect the enemy of good.
- Wowfunhappy 4y agoOkay, that's fair, there are a lot of symbols in unicode. But c'mon, it's pretty obvious how roman numeral unicode characters should be read. The software shouldn't choke on those!
- Joker_vD 4y ago> pretty obvious how roman numerals should be read If you know that they actually exist as separate code points then yeah, sure, it's kinda obvious. The same is true about "AV", "st" and "ſt" ligatures, I guess.
- usrbinbash 4y ago> But c'mon, it's pretty obvious how roman numerals should be read. Maybe it is in this special case, but it still sets a precedent that could lead to making something that should be simple and robust to something that is overengineered and complex, just because it has to accomodate an ever-increasing number of special cases, despite these cases making up a miniscule amount of the problem space. However, now that I think of it, it isn't even obvious in this special case. Yes roman numerals do have unicode code points, but many documents just represent them using latin letters instead, because it is easier to type and doesn't require unicode support. Now, should the screen-reader read "IV" as "Four" or as "I-Vee", which is a common abbreviation in medical texts? Sure, we could say "If it isn't using the unicode roman numeral, then its latin letters", and I am sure that it wouldn't be long for people complaining that this is a missing feature.
- Wowfunhappy 4y ago> Sure, we could say "If it isn't using the unicode roman numeral, then its latin letters". No, I don't think screen readers should do that unilaterally! That case—where latin letters are used—is legitimately tricky for the reasons you describe. If screen readers get it wrong, I understand. What I'm talking about is the other case, in which the author used the roman numeral unicode characters specifically. Most of the software in the article did worse with those characters compared to the latin letters! There's really no excuse for that!
- usrbinbash 4y ago> There's really no excuse for that! Yes there is, and it's a pretty good one: The occurrence of these symbols make up a miniscule fraction of the problem space, and they can be written using our everyday arabic numerals without any loss in context or information.
- Wowfunhappy 4y agoI'm not sure I understand. I do not think that any time anyone on the internet writes "XV" to represent "15", they should be required to use the special roman numeral characters. I certainly don't intend to do that in my own life—I don't have enough hours in my day—even though I recognize the problem it represents for screen readers. However, if I do go out of my way to use the special unicode characters, that should not trip up screen readers! I have done the screen reader an extra favor that they should be taking advantage of. As an imperfect analogy, imagine if a screen reader was ignoring the alt text on images. You might say "well, images with alt text represent a small fraction of the problem space", considering how many photos are casually uploaded to Instagram every day. And I agree—the internet is always going to contain some untagged images, and especially with the rise of AI, I'd love to see screen readers actually attempt to describe them. However, that is not an excuse for ignoring alt text when it is present! Covering "edge cases" like this is the core job of the screen reader!
- usrbinbash 4y ago
- deleted 4y ago[deleted]
- banannaise 4y agoRight, it's hard to make a screen reader understand context, but that's just a development/UX challenge. The point of a screen reader isn't to simply declare what characters are on the screen. Otherwise, it wouldn't be able to read English. The point of a screen reader is to interpret human-readable text into human-listenable sound. How is that character typically used? Perhaps as a scientific signifier of units? Or as an insert for another letter to create emphasis? Maybe both! Can the screen reader infer the context from the surrounding text? Or perhaps it's just a rarely used symbol, and the default output should be a blip that indicates a symbol or something.