// glossary entry

robots.txt

What is robots.txt?

The robots.txt file is a text file located in the root directory of a domain that provides search engine bots with rules for accessing directories and URLs. The "User-agent" and "Disallow" directives can be used to exclude areas from crawling, and the "Sitemap" directive can be used to reference the XML sitemap. Important: robots.txt prevents crawling, not indexing, the `noindex` tag in the page header is responsible for controlling indexing. At Waterproof Web Wizard, we check the robots.txt file during every onboarding: incorrectly configured “Disallow” rules are among the most common reasons for the question, “Why am I not ranking?”

// synonyms

  • robots.txt
  • Robots Exclusion Protocol
Is the term still unclear?

Let's talk for 15 minutes

If a term in this glossary isn't entirely clear in the context of your project, or if you want to know whether the concept is relevant to your situation-just ask us. The initial consultation is free and non-binding.

Free No obligation 15 minutes