{"id":25563,"date":"2026-09-05T02:45:42","date_gmt":"2026-09-04T23:45:42","guid":{"rendered":"https:\/\/alienroad.com\/google-bilgi-bankasi\/keep-redacted-information-out-of-google-search\/"},"modified":"2026-09-05T02:45:42","modified_gmt":"2026-09-04T23:45:42","slug":"keep-redacted-information-out-of-google-search","status":"publish","type":"ar_kb","link":"https:\/\/alienroad.com\/google-bilgi-bankasi\/keep-redacted-information-out-of-google-search\/","title":{"rendered":"Keep redacted information out of Google Search"},"content":{"rendered":"<p>\n      When publishing documents and images on the web, you may unintentionally publish information<br \/>\n      beyond what is immediately visible to the human eye. In particular, information that you might<br \/>\n      not see, or that was intended to be redacted,  might be included in some document formats and<br \/>\n      visible to search engines.\n    <\/p>\n<p>Because search engines index public material on the web, including images, content that is<br \/>\n      not completely redacted can potentially be findable in search engines. Assistive technologies<br \/>\n      like screen readers can make this seemingly &#8220;hidden&#8221; content more easily accessible, and<br \/>\n      common image understanding techniques like optical character recognition (OCR) similarly make<br \/>\n      it possible to search for this content.\n    <\/p>\n<p>Even though putting text in a tiny font, using a font color that&#8217;s the same as the background<br \/>\n      the text is on, or covering text with an image may make something invisible to the human eye,<br \/>\n      these methods don&#8217;t actually redact material in a way that prevents search engines from<br \/>\n      indexing it and making it findable.\n    <\/p>\n<p>\n      Similarly, some document types include information in various ways that aren&#8217;t immediately<br \/>\n      visible. They might include the document&#8217;s change history, allowing users to see text that has<br \/>\n      been redacted or altered. They might retain the full versions of images that contain cropped<br \/>\n      or redacted information. There might also be metadata that&#8217;s included in a file, which is not<br \/>\n      immediately visible, that may list the names of people who accessed or edited the file.\n    <\/p>\n<p>\n      All of this information can remain even when a document is exported or converted from one<br \/>\n      format to another. If you need to remove information from a file, it&#8217;s critical that the<br \/>\n      information is removed completely from the file before that file is made public.\n    <\/p>\n<p>\n      Here are some best practices for how to appropriately redact information from documents that<br \/>\n      you don&#8217;t want to be indexed and made discoverable via Google Search.\n    <\/p>\n<h2 id=\"edit-and-export-images-before-embedding-them\" tabindex=\"-1\">Edit and export images before embedding them<\/h2>\n<p>\n      Google Search lists images that it finds across the web, both those that are on web pages or<br \/>\n      those that are embedded into various document formats. Embedded images are sometimes edited<br \/>\n      using only the containing document&#8217;s editing tools. This can cause this redaction to fail when<br \/>\n      an image is indexed apart from the document. That is why it&#8217;s best to edit images before<br \/>\n      embedding them into a document, not after. In particular:\n    <\/p>\n<ul>\n<li>Crop out unwanted information from images before embedding them into documents. Some<br \/>\n        document editing tools (such as word processors or slide creation tools) will maintain any<br \/>\n        uncropped images that you use in the public version of the document, so be sure to review<br \/>\n        the tool&#8217;s documentation thoroughly.\n      <\/li>\n<li>\n        Completely remove or obscure any text or other non-public parts of the image, as OCR systems<br \/>\n        may turn any image text seen into searchable text.<\/li>\n<li>\n        Remove any undesired metadata.\n      <\/li>\n<\/ul>\n<p>\n      After following the suggestions in this document, export or save the updated images as non-vector or<br \/>\n      flattened image file formats such as PNG or WEBP. This prevents those parts of the images from<br \/>\n      being inadvertently included in a public document.\n    <\/p>\n<h2 id=\"edit-or-remove-unwanted-text-before-moving-to-a-public-file-format\" tabindex=\"-1\">Edit or remove unwanted text before moving to a public file format<\/h2>\n<p>\n      Before you generate the public document, remove any text that you don&#8217;t want displayed in the<br \/>\n      final version of the file. Move to a public format that does not keep your previous change<br \/>\n      history. Here are more specific tips:\n    <\/p>\n<ul>\n<li>Use proper document redacting tools if a file needs to have information redacted. For<br \/>\n        example, avoid placing black rectangles over text as a redaction method, as this can result<br \/>\n        in the text still being included in the public document.\n      <\/li>\n<li>\n        Double-check the document metadata in the public file.\n      <\/li>\n<li>\n        Follow the <a href=\"https:\/\/www.google.com\/search?q=document+redaction+best+practices\" class=\"external-link\">document redaction best practices<\/a><br \/>\n        for the format that you are using (PDF, image, etc).\n      <\/li>\n<li>\n        Consider information in the URL or file name itself. Even if a part of a website is<br \/>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/robots-txt-intro\/\">blocked by robots.txt<\/a>, the<br \/>\n        URLs may be indexed in search (without their content). Use hashes in URL parameters instead<br \/>\n        of email addresses or names.\n      <\/li>\n<li>\n        Consider using authentication to limit access to the redacted content. Serve the resulting<br \/>\n        login page with a<br \/>\n        <a href=\"https:\/\/alienroad.com\/google-bilgi-bankasi\/block-search-indexing-with-noindex\/\"><code>noindex<\/code> <span>robots<\/span> <code>meta<\/code> tag<\/a><br \/>\n        to block indexing.\n      <\/li>\n<li>\n        When publishing, make sure that the website is<br \/>\n        <a href=\"https:\/\/support.google.com\/webmasters\/answer\/9008080\" class=\"external-link\">verified in Google Search Console<\/a>.<br \/>\n        This allows quick removal action, if needed.\n      <\/li>\n<\/ul>\n<h2 id=\"what-to-do-if-unredacted-or-improperly-redacted-documents-are-indexed-in-search\" tabindex=\"-1\">What to do if unredacted or improperly redacted documents are indexed in Search<\/h2>\n<ol>\n<li>\n        Remove the live document from the website or location where you published it.\n      <\/li>\n<li>\n        Use the <a href=\"https:\/\/support.google.com\/webmasters\/answer\/9689846\" class=\"external-link\">Removals tool<\/a><br \/>\n        for the verified site to remove the documents in question from Search. Use a URL prefix if<br \/>\n        you need to remove many documents. For verified sites, a URL removal generally takes less<br \/>\n        than a day. This prevents the document in question from appearing for any searches for<br \/>\n        redacted content.\n      <\/li>\n<li>\n        Host the properly redacted document under a different URL. This makes sure that any newly<br \/>\n        indexed version is of the new document, and not an older version of the document (since<br \/>\n        recrawling of URLs and updating them in a search index can take a bit of time). Update any<br \/>\n        links to those documents.\n      <\/li>\n<li>\n        Contact any other site that may also be hosting the improperly redacted documents and ask<br \/>\n        them to take them down as well. Ask them to use the Removals tool in their Search Console<br \/>\n        account, or you can use the<br \/>\n        <a href=\"https:\/\/support.google.com\/webmasters\/answer\/7041154\" class=\"external-link\">Outdated Content tool<\/a><br \/>\n        to ask Google&#8217;s systems to update the search results.\n      <\/li>\n<li>\n        Allow the URL removal requests to expire (this happens after the URLs were either updated in<br \/>\n        the Google Search index, or after about 6 months).\n      <\/li>\n<\/ol>\n","protected":false},"excerpt":{"rendered":"<p>Learn best practices for how to redact information from documents that you don&#8217;t want to be indexed and made discoverable via Google Search.<\/p>\n","protected":false},"menu_order":49,"template":"","meta":{"footnotes":""},"ar_kb_kategori":[688],"ar_kb_etiket":[],"class_list":["post-25563","ar_kb","type-ar_kb","status-publish","has-post-thumbnail","hentry","ar_kb_kategori-removals-crawling-and-indexing"],"_links":{"self":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/25563","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb"}],"about":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/types\/ar_kb"}],"version-history":[{"count":0,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb\/25563\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media\/27583"}],"wp:attachment":[{"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/media?parent=25563"}],"wp:term":[{"taxonomy":"ar_kb_kategori","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_kategori?post=25563"},{"taxonomy":"ar_kb_etiket","embeddable":true,"href":"https:\/\/alienroad.com\/wp-json\/wp\/v2\/ar_kb_etiket?post=25563"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}