{"id":23907,"date":"2007-08-15T00:00:00","date_gmt":"2007-08-15T00:00:00","guid":{"rendered":"https:\/\/alienroad.com\/google-bilgi-bankasi\/new-robots-txt-feature-and-rep-meta-tags\/"},"modified":"2007-08-15T00:00:00","modified_gmt":"2007-08-15T00:00:00","slug":"new-robots-txt-feature-and-rep-meta-tags","status":"publish","type":"ar_kb","link":"https:\/\/alienroad.com\/google-bilgi-bankasi\/new-robots-txt-feature-and-rep-meta-tags\/","title":{"rendered":"New robots.txt feature and REP Meta Tags"},"content":{"rendered":"<p class=\"gargardate\">Wednesday, August 15, 2007<\/p>\n<p>\n  We&#8217;ve improved Webmaster Central&#8217;s robots.txt analysis tool to recognize Sitemap declarations and<br \/>\n  relative URLs. Earlier versions weren&#8217;t aware of Sitemaps at all, and understood only absolute<br \/>\n  URLs; anything else was reported as <code>Syntax not understood<\/code>. The improved version now<br \/>\n  tells you whether your Sitemap&#8217;s URL and scope are valid. You can also test against relative<br \/>\n  URLs with a lot less typing.\n<\/p>\n<p>\n  Reporting is better, too. You&#8217;ll now be told of multiple problems per line if they exist, unlike<br \/>\n  earlier versions which only reported the first problem encountered. And we&#8217;ve made other general<br \/>\n  improvements to analysis and validation.\n<\/p>\n<p>\n  Imagine that you&#8217;re responsible for the domain <b>www.example.com<\/b> and you want search engines<br \/>\n  to index everything on your site, except for your <code>\/images<\/code> folder. You also want to<br \/>\n  make sure your Sitemap gets noticed, so you save the following as your robots.txt file:\n<\/p>\n<div><\/div>\n<p>\n  You visit Webmaster Central to test your site against the robots.txt analysis tool using these two<br \/>\n  test URLs:\n<\/p>\n<div><\/div>\n<p>Earlier versions of the tool would have reported this:<\/p>\n<p><img decoding=\"async\" alt=\"robots.txt tester report in webmaster tools pre improvements\" src=\"https:\/\/alienroad.com\/wp-content\/uploads\/kb-gorsel\/g-9252e0adb3be.png\" loading=\"lazy\"><\/p>\n<p>The improved version tells you more about that robots.txt file:<\/p>\n<p><img decoding=\"async\" alt=\"robots.txt tester report in webmaster tools pre improvements\" src=\"https:\/\/alienroad.com\/wp-content\/uploads\/kb-gorsel\/g-d3e8f8b55966.png\" loading=\"lazy\"><\/p>\n<p>\n  See for yourself at<br \/>\n  <a href=\"https:\/\/search.google.com\/search-console\" class=\"external-link\">https:\/\/www.google.com\/webmasters\/tools<\/a>.\n<\/p>\n<p>\n  We also want to make sure you&#8217;ve heard about the new <code>unavailable_after<\/code> <code>meta<\/code> tag<br \/>\n  announced by Dan Crow on the<br \/>\n  <a href=\"https:\/\/googleblog.blogspot.com\/2007\/07\/robots-exclusion-protocol-now-with-even.html\" class=\"external-link\">Official Google Blog<\/a><br \/>\n  a few weeks ago. This allows for a more dynamic relationship between your site and Googlebot. Just<br \/>\n  think, with <b>www.example.com<\/b>, any time you have a temporarily available news story or<br \/>\n  limited offer sale or promotion page, you can specify the exact date and time you want specific<br \/>\n  pages to stop being crawled and indexed.\n<\/p>\n<p>\n  Let&#8217;s assume you&#8217;re running a promotion that expires at the end of 2007. In the headers of page<br \/>\n  <b>www.example.com\/2007promotion.html<\/b>, you would use the following:\n<\/p>\n<div><\/div>\n<p>\n  The second exciting news: the new <code>X-Robots-Tag<\/code> rule, which adds<br \/>\n  <a href=\"https:\/\/googleblog.blogspot.com\/2007\/02\/robots-exclusion-protocol.html\" class=\"external-link\">Robots Exclusion Protocol<\/a><br \/>\n  (REP) <code>meta<\/code> tag support for non-HTML pages! Finally, you can have the same control<br \/>\n  over your videos, spreadsheets, and other indexed file types. Using the example above, let&#8217;s say<br \/>\n  your promotion page is in PDF format. For <b>www.example.com\/2007promotion.pdf<\/b>, you would use<br \/>\n  the following:\n<\/p>\n<div><\/div>\n<p>\n  Remember, REP <code>meta<\/code> tags can be useful for implementing noarchive, nosnippet, and now<br \/>\n  <code>unavailable_after<\/code> tags for page-level instruction, as opposed to robots.txt, which<br \/>\n  is controlled at the domain root. We get requests from bloggers and webmasters for these features,<br \/>\n  so enjoy. If you have other suggestions, keep them coming. Any questions? Please ask them in the<br \/>\n  <a href=\"https:\/\/support.google.com\/webmasters\/community\" class=\"external-link\">Webmaster Help Group<\/a>.\n<\/p>\n<p class=\"byline-author\">\n  Posted by John Blackburn, Webmaster Tools and Matt Dougherty, Search Quality<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Wednesday, August 15, 2007 We&#8217;ve improved Webmaster Central&#8217;s robots.txt analysis tool to recognize Sitemap declarations and relative URLs. Earlier versions weren&#8217;t aware of Sitemaps at all, and understood only absolute URLs; anything else was reported as Syntax not understood. The improved version now tells you whether your Sitemap&#8217;s URL and scope are valid. You can [&hellip;]<\/p>\n","protected":false},"menu_order":86260,"template":"","meta":{"footnotes":""},"ar_kb_kategori":[665],"ar_kb_etiket":[],"class_list":["post-23907","ar_kb","type-ar_kb","status-publish","has-post-thumbnail","hentry","ar_kb_kategori-blog"],"_links":{"self":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/23907","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb"}],"about":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/types\/ar_kb"}],"version-history":[{"count":0,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/23907\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media\/26534"}],"wp:attachment":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media?parent=23907"}],"wp:term":[{"taxonomy":"ar_kb_kategori","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_kategori?post=23907"},{"taxonomy":"ar_kb_etiket","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_etiket?post=23907"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}