3 ms·
I dont know if this is naive, but wouldnt the data model/storage strategy of the index be influenced by the ranking algorithms that use it? If thats the case, t
by redditmigrant 15y ago
I dont know if this is naive, but wouldnt the data model/storage strategy of the index be influenced by the ranking algorithms that use it? If thats the case, then I would presume google's index tries to store the data in a form thats efficient for their ranking algorithm to work off of and it might not be in the best format for say bing/yahoo to use.
- chrislomax 15y agoI think they are referring to the data in its rawest format before they have indexed and ranked the information themselves. They will all crawl the information in exactly the same way. They will just take the plain text and store it. I don't think any bot would actually do anything else with the data on the fly. If you think about it, it does make sense in a lot of respects. I have dealt with a lot of companies that sell data, the only difference is this data is freely available to everyone so everyone thinks they should crawl the information themselves. The only people who lose out are the people paying the bandwidth bills. The internet would actually be slower due to the amount of information passing around when it is not needed. This idea makes more sense the more we discuss it
- Lewisham 15y agoYes, it would be fair to assume that index optimization is also part of the Secret Sauce, unless you store the raw data. Storing the raw data also requires Secret Sauce like Google FIle System, and you'll end up with the sarcastic comment above that the Internet is the raw data and we're back at square one.