{"id":24028,"date":"2008-09-12T00:00:00","date_gmt":"2008-09-12T00:00:00","guid":{"rendered":"https:\/\/alienroad.com\/google-bilgi-bankasi\/demystifying-the-duplicate-content-penalty\/"},"modified":"2008-09-12T00:00:00","modified_gmt":"2008-09-12T00:00:00","slug":"demystifying-the-duplicate-content-penalty","status":"publish","type":"ar_kb","link":"https:\/\/alienroad.com\/google-bilgi-bankasi\/demystifying-the-duplicate-content-penalty\/","title":{"rendered":"Demystifying the &#8220;duplicate content penalty&#8221;"},"content":{"rendered":"<p class=\"gargardate\">Friday, September 12, 2008<\/p>\n<p>\n  Duplicate content. There&#8217;s just something about it. We<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/deftly-dealing-with-duplicate-content\/\">keep<\/a><br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/duplicate-content-summit-at-smx-advanced\/\">writing<\/a><br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/google-duplicate-content-caused-by-url-parameters-and-you\/\">about<\/a><br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/duplicate-content-due-to-scrapers\/\">it<\/a>, and people keep asking<br \/>\n  about it. In particular, I still hear a lot of webmasters worrying about whether they may have a<br \/>\n  &#8220;duplicate content penalty.&#8221;\n<\/p>\n<p>\n  Let&#8217;s put this to bed once and for all, folks: There&#8217;s no such thing as a &#8220;duplicate content<br \/>\n  penalty.&#8221; At least, not in the way most people mean when they say that.\n<\/p>\n<p>\n  There are some penalties that are related to the idea of having the same content as another<br \/>\n  site&mdash;for example, if you&#8217;re scraping content from other sites and republishing it, or if<br \/>\n  you republish content without adding any additional value. These tactics are clearly outlined<br \/>\n  (and discouraged) in our<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/search-essentials-overview\/\">Webmaster Guidelines<\/a>:\n<\/p>\n<ul>\n<li>\n    Don&#8217;t create multiple pages, subdomains, or domains with substantially<br \/>\n    <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/specify-canonical\/\">duplicate content<\/a>.\n  <\/li>\n<li>\n    Avoid&#8230; &#8220;cookie cutter&#8221; approaches such as affiliate programs with<br \/>\n    <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/spam-policies\/\">little or no original content<\/a>.\n  <\/li>\n<li>\n    If your site participates in an<br \/>\n    <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/spam-policies\/#thin-affiliate-pages\">affiliate program<\/a>, make sure<br \/>\n    that your site adds value. Provide unique and relevant content that gives users a reason to<br \/>\n    visit your site first.\n  <\/li>\n<\/ul>\n<p>\n  (Note that while scraping content from others is discouraged, having others scrape you is a<br \/>\n  different story;<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/duplicate-content-due-to-scrapers\/\">check out this post<\/a><br \/>\n  if you&#8217;re worried about being scraped.)\n<\/p>\n<p>\n  But most site owners whom I hear worrying about duplicate content aren&#8217;t talking about scraping or<br \/>\n  domain farms; they&#8217;re talking about things like having multiple URLs on the same domain that point<br \/>\n  to the same content. Like<br \/>\n  <code>www.example.com\/skates.asp?color=black&amp;brand=riedell<\/code><br \/>\n  and <code>www.example.com\/skates.asp?brand=riedell&amp;color=black<\/code>. Having this type of<br \/>\n  duplicate content on your site can potentially affect your site&#8217;s performance, but it doesn&#8217;t<br \/>\n  cause penalties. From our article on<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/specify-canonical\/\">duplicate content<\/a>:\n<\/p>\n<p>\n  Duplicate content on a site is not grounds for action on that site unless it appears that the<br \/>\n  intent of the duplicate content is to be deceptive and manipulate search engine results. If your<br \/>\n  site suffers from duplicate content issues, and you don&#8217;t follow the advice listed above, we do<br \/>\n  a good job of choosing a version of the content to show in our search results.\n<\/p>\n<p>\n  This type of non-malicious duplication is fairly common, especially since many CMSs don&#8217;t handle<br \/>\n  this well by default. So when people say that having this type of duplicate content can affect<br \/>\n  your site, it&#8217;s not because you&#8217;re likely to be penalized; it&#8217;s simply due to the way that web<br \/>\n  sites and search engines work.\n<\/p>\n<p>\n  Most search engines strive for a certain level of variety; they want to show you ten different<br \/>\n  results on a search results page, not ten different URLs that all have the same content. To this<br \/>\n  end, Google tries to filter out duplicate documents so that users experience less redundancy. You<br \/>\n  can find details in<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/google-duplicate-content-caused-by-url-parameters-and-you\/\">this blog post<\/a>, which<br \/>\n  states:\n<\/p>\n<ol>\n<li>\n    When we detect duplicate content, such as through variations caused by URL parameters, we group<br \/>\n    the duplicate URLs into one cluster.\n  <\/li>\n<li>We select what we think is the &#8220;best&#8221; URL to represent the cluster in search results.<\/li>\n<li>\n    We then consolidate properties of the URLs in the cluster, such as link popularity, to the<br \/>\n    representative URL.\n  <\/li>\n<\/ol>\n<p>Here&#8217;s how this could affect you as a webmaster:<\/p>\n<ul>\n<li>\n    In step 2, Google&#8217;s idea of what the &#8220;best&#8221; URL is might not be the same as your idea. If you<br \/>\n    want to have control over whether<br \/>\n    <code>www.example.com\/skates.asp?color=black&amp;brand=riedell<\/code> ou<br \/>\n    <code>www.example.com\/skates.asp?brand=riedell&amp;color=black<\/code> gets shown in our search<br \/>\n    results, you may want to take action to mitigate your duplication.<br \/>\n    One way of letting us know which URL you prefer is by including the preferred URL in your<br \/>\n    <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/sitemaps-overview\/\">Sitemap<\/a>.\n  <\/li>\n<li>\n    In step 3, if we aren&#8217;t able to detect all the duplicates of a particular page, we won&#8217;t be<br \/>\n    able to consolidate all of their properties. This may dilute the strength of that content&#8217;s<br \/>\n    ranking signals by splitting them across multiple URLs.\n  <\/li>\n<\/ul>\n<p>\n  In most cases Google does a good job of handling this type of duplication. However, you may also<br \/>\n  want to consider content that&#8217;s being duplicated across domains. In particular, deciding to build<br \/>\n  a site whose purpose inherently involves content duplication is something you should think twice<br \/>\n  about if your business model is going to rely on search traffic, unless you can add a lot of<br \/>\n  additional value for users. For example, we sometimes hear from Amazon.com affiliates who are<br \/>\n  having a hard time ranking for content that originates solely from Amazon. Is this because Google<br \/>\n  wants to stop them from trying to sell<br \/>\n  <a href=\"https:\/\/www.amazon.com\/Everyone-Poops-My-Body-Science\/dp\/0916291456\" class=\"external-link\">Everyone Poops<\/a>?<br \/>\n  No; it&#8217;s because <em>how the heck are they going to outrank Amazon<\/em> if they&#8217;re providing<br \/>\n  the exact same listing? Amazon has a lot of online business authority (most likely more than<br \/>\n  a typical Amazon affiliate site does), and the average Google search user probably wants the<br \/>\n  original information on Amazon, unless the affiliate site has added a significant amount of<br \/>\n  <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/spam-policies\/#thin-affiliate-pages\">additional value<\/a>.\n<\/p>\n<p>\n  Lastly, consider the effect that duplication can have on your site&#8217;s bandwidth. Duplicated<br \/>\n  content can lead to inefficient crawling: when Googlebot discovers ten URLs on your site, it has<br \/>\n  to crawl each of those URLs before it knows whether they contain the same content (and thus before<br \/>\n  we can group them as described above). The more time and resources that Googlebot spends crawling<br \/>\n  duplicate content across multiple URLs, the less time it has to get to the rest of your content.\n<\/p>\n<p>\n  In summary: Having duplicate content can affect your site in a variety of ways; but unless you&#8217;ve<br \/>\n  been duplicating deliberately, it&#8217;s unlikely that one of those ways will be a penalty. This means<br \/>\n  that:\n<\/p>\n<ul>\n<li>\n    You typically don&#8217;t need to submit a reconsideration request when you&#8217;re cleaning up innocently<br \/>\n    duplicated content.\n  <\/li>\n<li>\n    If you&#8217;re a webmaster of beginner-to-intermediate savviness, you probably don&#8217;t need to put<br \/>\n    too much energy into worrying about duplicate content, since most search engines have ways of<br \/>\n    handling it.\n  <\/li>\n<li>\n    You can help your fellow webmasters by not perpetuating the myth of duplicate content penalties!<br \/>\n    The remedies for duplicate content are entirely within your control. Here are some good places<br \/>\n    to start:<\/p>\n<ul>\n<li>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/specify-canonical\/\">Avoid creating duplicate content<\/a>\n      <\/li>\n<li>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/deftly-dealing-with-duplicate-content\/\">Deftly dealing with duplicate content<\/a>\n      <\/li>\n<li>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/duplicate-content-summit-at-smx-advanced\/\">Duplicate content summit at SMX Advanced<\/a>\n      <\/li>\n<li>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/google-duplicate-content-caused-by-url-parameters-and-you\/\">Google, duplicate content caused by URL parameters, and you<\/a>\n      <\/li>\n<\/ul>\n<\/li>\n<\/ul>\n<p class=\"byline-author\">Posted by Susan Moskwa, Webmaster Trends Analyst<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Friday, September 12, 2008 Duplicate content. There&#8217;s just something about it. We keep writing about it, and people keep asking about it. In particular, I still hear a lot of webmasters worrying about whether they may have a &#8220;duplicate content penalty.&#8221; Let&#8217;s put this to bed once and for all, folks: There&#8217;s no such thing [&hellip;]<\/p>\n","protected":false},"menu_order":85866,"template":"","meta":{"footnotes":""},"ar_kb_kategori":[665],"ar_kb_etiket":[],"class_list":["post-24028","ar_kb","type-ar_kb","status-publish","has-post-thumbnail","hentry","ar_kb_kategori-blog"],"_links":{"self":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/24028","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb"}],"about":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/types\/ar_kb"}],"version-history":[{"count":0,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/24028\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media\/26628"}],"wp:attachment":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media?parent=24028"}],"wp:term":[{"taxonomy":"ar_kb_kategori","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_kategori?post=24028"},{"taxonomy":"ar_kb_etiket","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_etiket?post=24028"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}