HTTP Header Checker

Inspect the response headers behind any URL. The SEO-relevant ones are called out first, including the X-Robots-Tag that can keep a page out of Google without leaving any trace in the HTML.

Paste a full URL to check that page, or a bare domain to check its home page. Every response header comes back, and the ones that decide how the page is indexed, cached, and compressed are called out first with what each value means.

When a page disappears and the HTML looks fine

The usual diagnosis path for a deindexed page runs through the meta robots tag, the canonical, and robots.txt. When all three come back clean and the page is still gone, the answer is almost always in the response headers, because a header can carry the same directives and nothing in the page source reveals it.

This happens most often when a rule was added at the CDN or load balancer rather than in the application, which is a common way to keep a staging environment out of search. The rule outlives the environment, and a path pattern that once matched only staging starts matching production paths after a restructure.

An X-Robots-Tag is read one line at a time, and a line may name the crawler it addresses before its directives. A response carrying all on one line and otherbot: noindex on the next is indexed by Google normally, so this tool reports a noindex as blocking only where a Google crawler reads it, and names the crawler where one is named.

The Link header is worth the same attention. A canonical declared there competes with the one in the HTML, and when they disagree the outcome is not something you can reason about from the page alone.

One thing this check cannot do for you is arrive as Googlebot. A rule at the CDN may serve a header to Google and to nobody else, which is exactly the rule that hides best. When the headers here look right and the page is still missing, the URL Inspection tool in Search Console fetches it as Googlebot and settles the question.

Frequently Asked Questions

What is an X-Robots-Tag and why does it matter?

It is a way of sending robots directives in the HTTP response rather than in the HTML. A header carrying noindex removes a page from search exactly as a meta tag would, but nothing in the page source shows it, so it is the hardest deindexing cause to find. If a page has vanished from Google and its HTML looks correct, check this header first.

Get Google and ChatGPT traffic on autopilot.

Start today and generate your first article within 15 minutes.

No credit card required
Content Plan