2 ms·
> In so much as a solution depends on "cleverness" rather than raw computing power, it'll have room for competition. Unfortunately, it's looking more and more
by yid 11y ago
> In so much as a solution depends on "cleverness" rather than raw computing power, it'll have room for competition.
Unfortunately, it's looking more and more like it's going to be a competition on training data and raw computational power, and it's hard to compete with Google's corpora from the web, gmail, captchas, maps, etc. -- not to mention Google's tremendous number crunching resources.
- jjackson21 11y agoWhat other data sets exist that would be on par with Google's web/gmail/maps/etc. data sets for training models? Is there a chance Google would every make a version of their data sets available to the public?
- singhrac 11y agoAt least right now there aren't many other places that have as much data flowing through them, or at least as much almost unrestricted access to data (for example, Amazon has plenty of data but very little legal access). And possibly, though for privacy reasons in a much more filtered down form (see the word2vec Google News dataset). I find it unlikely for private data though.