3 ms·
Yes, thanks for your interest - we only have 50 concurrent users licensed. Never thought we'd get this much interest. :)
by mcfrank 10y ago
Yes, thanks for your interest - we only have 50 concurrent users licensed. Never thought we'd get this much interest. :)
- rspeer 10y agoWould it be possible to mirror just the data somewhere else, such as S3? I don't need the R code, but this sounds like it would make good companion data to my own wordfreq [1]. It would be interesting to see which words are learned early but relatively uncommon in corpora, and generally to be able to measure differences in register between child and adult language. [1] https://github.com/LuminosoInsight/wordfreq https://github.com/LuminosoInsight/wordfreq
- mikabr 10y agowe have an R package which you can use to access the data: https://github.com/langcog/wordbankr https://github.com/langcog/wordbankr we've done some analyses predicting words' learnability from frequency and other factors: http://langcog.stanford.edu/papers_new/braginsky-2016-underrev.pdf http://langcog.stanford.edu/papers_new/braginsky-2016-underr...
- mcfrank 10y agoVery cool! All our code is at http://github.com/langcog/wordbank http://github.com/langcog/wordbank and you can access the database directly using the wordbankr R package (on cran). A paper doing something similar to what you describe is in prep, with a conference version here: http://langcog.stanford.edu/papers_new/braginsky-2016-underrev.pdf http://langcog.stanford.edu/papers_new/braginsky-2016-underr...