robots.txt · headers · meta · sitemaps

Ask the website
before you crawl

Instantly analyse robots.txt, crawl directives, AI bot permissions, meta robots, sitemaps and HTTP headers to understand how a website publishes its crawling preferences.

Try

No signup required • Free • Instant analysis

Why SiteAsker

Everything a site says about crawling, in one place

See the published preferences

Read what a site actually states about crawling, instead of guessing from its footer.

Inspect robots.txt

Every group, every rule, parsed the way a real crawler resolves them.

Check AI crawler rules

GPTBot, ClaudeBot, Google-Extended, CCBot, PerplexityBot and Applebot, each read separately.

Find the sitemaps

Listed sitemaps, plus a look at the usual address when none are declared.

Read the HTTP headers

X-Robots-Tag directives travel in headers and never appear in robots.txt.

Read the meta tags

index, noindex, follow, nofollow — taken straight from the homepage HTML.

Built for developers

One request, one structured report, and the raw exchange that produced it.

Free to use

Sign in once and check as many sites as you need.

How it works

Three steps, about four seconds

01

Enter a website

Paste any URL. A bare domain is enough — we work out the rest.

02

We ask the website

We request robots.txt, follow the redirects, read the response headers and the homepage meta tags.

03

Read the report

A single verdict with a confidence rating, and every rule we found behind it.

FAQ

Questions worth asking first

A plain text file at the root of a site that tells automated clients which paths the site would prefer they didn't request. It's a published preference, honoured by convention.