3 ms·
There is a difference between robots.txt blocking a page and noindexing a page. Blocking in robots.txt will stop Googlebot downloading that page and looking at
by TomAnthony 7y ago
There is a difference between robots.txt blocking a page and noindexing a page.
Blocking in robots.txt will stop Googlebot downloading that page and looking at the contents, but the page may still make it into the index on the basis of links to that page making it seem relevant (it will appear in the search results without a description snippet and will include a note about why).
To have a page not appear in the index you need to use a 'noindex' directive [1] either in the file itself or in the HTTP headers. However, if the file is blocked in robots.txt then note Google cannot read that noindex directive.
Also, in the StackOverflow response you linked to that the user agent is listed just as 'Google', but it should be 'Googlebot' as per the 'User agent token (product token)' table column listed in [2].
Good luck! :)
[1] https://support.google.com/webmasters/answer/93710?hl=en https://support.google.com/webmasters/answer/93710?hl=en
[2] https://support.google.com/webmasters/answer/1061943 https://support.google.com/webmasters/answer/1061943