2 ms·
Ha! Some very nice examples, I have to say :) Anyway, You’re completely right. Italian is `und` due to LTE 10 characters, the others are slightly off due to sh
by wooorm 12y ago
Ha! Some very nice examples, I have to say :)
Anyway, You’re completely right. Italian is `und` due to LTE 10 characters, the others are slightly off due to short input too, but the demo (http://wooorm.github.io/franc/ http://wooorm.github.io/franc/) shows the correct languages in the second or third place though!
- jodent 12y agoNo it doesn't, still takes French for Catalan (French only comes at third place, after Italian), and Swedish for Dutch. (Arguably those are close languages, but hey, this is why I'm using this, right?)
- wooorm 12y agoBy `correct language` I mean the language you expect, by `second` and `third` I mean `2.` and `3.` in the previously mentioned demo: http://wooorm.github.io/franc/ http://wooorm.github.io/franc/). I think we’re talking about the same thing! Anyway, yeah, franc is for language detecting, but it’s optimised for many languages and works best at longer text. It’s a trade-off. For less languages and shorter texts, check out https://github.com/shuyo/ldig https://github.com/shuyo/ldig