7 ms·
The PageRank Citation Ranking: Bringing Order to the Web (1998) [pdf]
- androidfox 10y agoThis is 1998 Link. Is this still at least partially valid
- upen 10y agoI think Google stopped updating Pagerank two years back
- rbinv 10y agoYou're thinking of the "public" PageRank value previously displayed in Google's browser toolbars. Those are indeed no longer updated. However, there is still some variant of "PageRank" in use internally.
- visarga 10y agoThey are probably still using it, but I think they apply spam and duplicate content detection first and also limit PageRank flow to sites of similar topic. Another thing they could do is to nullify the old links for resold websites. Besides that, time spent on page and social signals such as likes and tweets probably count for more than links. A crappy spam site would be disregarded pretty soon, because people would immediately bounce back to Google and try other sites.
- rbinv 10y agoProbably somewhat. However, "link equity/power" has definitely (in part) been replaced by other ranking signals.
- amelius 10y agoYes, is Google still using this internally? If not, then what would be a more up to date reference?
- harigov 10y agoDidn't they move towards "RankBrain" which I believe is a ML based search engine. Links maybe one of those signals but I believe user activity maybe lot more interesting for them.
- andrewstuart2 10y agoAlso fun and related: http://infolab.stanford.edu/~sergey/ http://infolab.stanford.edu/~sergey/ And another paper by Page and Brin: http://infolab.stanford.edu/~backrub/google.html http://infolab.stanford.edu/~backrub/google.html
- andreygrehov 10y ago10+ years ago I needed to contact someone from Google and I found this page. At the time, I didn't know who Sergey Brin is. So I mailed the guy to an email provided at the bottom of the page and told my colleagues something like: I found a guy who works at Google, we should be all good, I just emailed him, let's see what he says. Everyone was laughing their heads off when they asked me the guy's name.
- gigatexal 10y agothe best part of the paper was this: "To test the utility of PageRank for search we built a web search engine called Go ogle" now a company worth hundreds of billions.
- hbbio 10y agoI love how the (now famous) researchers dismiss at the time what has become Google's main chase: "These types of personalized PageRanks are virtually immune to manipulation by commercial interests. For a page to get a high PageRank, it must convince an important page, or a lot of non-important pages to link to it. At worst, you can have manipulation in the form of buying advertisements(links) on important sites. But, this seems well under control since it costs money"
- rahrahrah 10y agoAnd yet, everything in that paragraph is absolutely true. If anything, they have stopped short from taking it to its logical conclusions. Since it costs money, only when more money than it costs is involved will this take place.
- rahrahrah 10y agoTo go further, the implicit assumption that they were making when stopping short was: Google will never be big enough that people will spend money to manipulate it. But it did, and people do. To me this reinforces that you should take seriously things like 50% bitcoin attacks. If bitcoins become valuable enough, someone will do it.
- karakal 10y agoIs it still possible in today's Academia setting to conceive a company like Google without oweing anything to the institution in which it was created? It looks like the first version of Google was even hosted on Stanford's computers.
- moultano 10y agoGoogle paid licencing fees to Stanford for pagerank in the form of stock. http://www.ipnav.com/blog/google-algorithm-earns-337-million-for-stanford/ http://www.ipnav.com/blog/google-algorithm-earns-337-million...
- jldugger 10y agoIt wasn't possible then either.
- Keverw 10y agoDo colleges own startups people create while going to school? If so that sounds horrible considering how much college costs. I heard if you work for a company while also developing a startup depending on the contract the company can claim they own the startup - even if you were working on it at home with your computer you bought with your own money and off the clock for them.
- ucaetano 10y agoNo, but they do own patents created by their researchers. Stanford didn't own Google, they owned the PageRank patent, which they licensed to Google in return for equity.
- Keverw 10y agoInteresting. "patents created by their researchers" so if it's a official school project then, and not someone doing it on their own? That would makes since. I don't get why schools need patents in the first place though.
- 10y ago
- ucaetano 10y agoAlso relevant: https://groups.google.com/forum/#!msg/comp.lang.java/aSPAJO05LIU/ushhUIQQ-ogJ https://groups.google.com/forum/#!msg/comp.lang.java/aSPAJO0...
- scoot 10y agoIn the mobile app I'm using the company name in your comment has split across two lines as "Go-ogle", which suddenly made so much sense! I was slightly disappointed that this isn't what it says in the paper, or presumably your comment! (Still no Ida why it was split when words are usually wrapped.)
- Xeoncross 10y agoAre there any other good reads on how this technology is implemented, the shortcomings since it was first released, improvements in the algorithm, splitting the calculations up into a map-reduce for the modern (larger) web, and anything else that might be a good read? I just googled for some implementations like this one in Go: https://github.com/dcadenas/pagerank https://github.com/dcadenas/pagerank
- combatentropy 10y agohttps://scholar.google.com/scholar?q=sergey+brin https://scholar.google.com/scholar?q=sergey+brin
- jzl 10y agoAnd here's the original patent. I never knew until now that Sergey Brin wasn't on it. Will expire in about a year and three months! https://www.google.com/patents/US6285999 https://www.google.com/patents/US6285999
- sparky_z 10y agoWell, they didn't call it Brinrank :)
- teddyknox 10y agoThis reminds me of the new Black Mirror season 3 pilot
- jacquesm 10y agoAny act of measuring a system changes that system. Pagerank did more to destroy the value of the web of links than any other technology before it because it was so good at measuring its value. Extracting that value then instantly leads to diminishing it because others (in this case the link spammers) want a slice of that huge pie. I believe that each and every technology that successfully manages to index the web in a new and useful way will further diminish the value of the web.
- codeulike 10y agoAny act of measuring a system changes that system. How about when I measure a stick using a ruler?
- nothrabannosir 10y agoHe means a system that can react to the measurements. Like how predicting the stock market is fundamentally impossible because whatever you do, the stock market can react to your prediction.
- codeulike 10y agoBut he said 'Any'. I'm looking for a counterexanple.
- ssn 10y agoBut does it work in practice? It doesn't seem to. PageRank is mostly a marketing tool. https://www.microsoft.com/en-us/research/publication/hits-on-the-web-how-does-it-compare/ https://www.microsoft.com/en-us/research/publication/hits-on... "The fact that in-degree features outperform PageRank under all measures is quite surprising. A possible explanation is that link-spammers have been targeting the published PageRank algorithm for many years, and that this has led to anomalies in the web graph that affect PageRank."
- srean 10y agoWell, if you are looking for a measure that is the easiest to game, abuse and spam, you cant go wrong with your pagerank killer: in-degree. It works less well than it used to, but its never used in isolation. Used in isolation its a pretty good porn detector.