3 ms·
> Proprietary algorithms can help, but they are secondary in importance to the data sets themselves. This is flat wrong. No algo, no intelligence or synthetic
by meeper16 10y ago
> Proprietary algorithms can help, but they are secondary in importance to the data sets themselves.
This is flat wrong. No algo, no intelligence or synthetic congnition
> The dramatic rise of Google provides a glimpse into what this kind of privileged access can enable. What allowed Google to rapidly take over the search market was not primarily its PageRank algorithm or clean interface, but these factors in combination with its early access to the data sets of AOL and Yahoo, which enabled it to train PageRank on the best available data on the planet and become substantially better at determining search relevance than any other product.
This is so wrong on so many fronts. A) They has open access to the web just like DMOZ and Yahoo and many others via crawling systems. B) They were attractive to software engineers who in turn made their IT depts in large corps switch to Google as the default search engine C) They stole the ad matching algo from Bill Gross which in turn made them successful. E) too many other factors to list
Lets also not forget that PageRank was a variant of an early link analysis algo that had been around a while.
- izendejas 10y agoOn the first statement, I think it's mostly right and he's merely just re-stating what Peter Norvig has stated all along albeit not as precisely. On the second statement, I believe you both missed the point. Google was able to acquire click-through data which helped train it's rankin algorithms better by establishing the partnerships with Yahoo and AOL. It was the implicit feedback, not the crawled data that helped. But that obviously wasn't the only factor--having very smart engineers/scientists matters. But I think now that a lot more is open sourced, computing power is cheaper, etc the OPs thesis is that data is now the hardest thing to come by, and on this I fully agree.