Gleanzy
Free tool · Domain and website

Robots.txt checker

Read a site's robots.txt and see which crawlers may visit which paths.

  • Rules grouped by user agent
  • Sitemap references
  • Parse problems flagged
Robots.txt checker20 free checks a day · no account

Give it something to look up — Glee will do the digging.

What this checks

How the robots.txt checker works

robots.txt is the file at the root of a site that tells crawlers what not to fetch. One wrong line — a stray Disallow: / left over from a staging site — can remove a whole site from search results, and it is one of the first things to check when traffic drops.

This checker fetches /robots.txt, parses it into groups by user agent, and lists the allow and disallow rules, crawl delays and sitemap references it declares, with any lines a crawler would ignore.

Reading the result

What each field means

Rules
Allow and Disallow lines per user agent.
Sitemaps
Sitemap files the site declares.
Interpretation limits

Where this check stops

Use it from your code

The same check, through the API

Same calculation, same answer, with a key. 5 credits per check ($0.50 per 1,000). A free account includes 1,000 credits a month.

Get a free API key

POST /v1/web {"url":"acme.com"}
Host: gleanzy.com
Authorization: Bearer $GLEANZY_KEY
Questions

Frequently asked

Does Disallow remove a page from search?

No — it stops crawling. Use a noindex directive on the page to keep it out.

Related checks

Ask the next question

One question answered

Put the whole company in context.

Start free