3 ms·
Speaking of rarely updated - version two will come out mid-July at the latest. I'm also going to attempt to attribute the sources to given countries or at leas
by berzerk0 9y ago
Speaking of rarely updated - version two will come out mid-July at the latest.
I'm also going to attempt to attribute the sources to given countries or at least language families. The list as it stands has a heavily US-centric bias.
I plan to grep through it and pull out things like Cyrillic or Arabic characters, and then repeat the process using only those characters. That's not to say that everyone who uses Arabic or Cyrillic characters is guaranteed to use them, but it's more specific.
If "password" is #2 in the world, perhaps пароль (password) will be the #2 most popular password with Russian Cyrillic Characters
- lucasgonze 9y agoI wonder about how to package your data as developer tools. There should be libraries in any language used to develop systems which validate user-entered passwords.
- berzerk0 9y agoDropbox has their zxcvbn method which incorporates a list of 30k of the most common passwords https://github.com/dropbox/zxcvbn https://github.com/dropbox/zxcvbn
- bradleyjg 9y agoIt looks like they gzip the file over the wire and then work with it uncompressed. That's probably a reasonable choice given they use only 30,000 passwords.