noindex and X-Robots-Tag on source pages
Published · By IndexChex
A noindex rule on the page that hosts a backlink tells Google not to show that page in search results. The link still exists in the HTML, but a page kept out of the index carries little practical value, so backlink monitors check both the robots meta tag and the X-Robots-Tag HTTP header on every source page.
What the directive does
Google documents two ways to keep a page out of its results: a noindex value in a robots <meta> tag in the page head, or the same value sent in an X-Robots-Tag HTTP response header. When Googlebot next fetches the page and sees either one, Google drops the page from search results, regardless of how many other sites link to it. A <meta name="googlebot"> tag applies the same rule to Google only.
For a backlink owner this matters because the link lives on someone else's page. The anchor text and href can be unchanged while the page around it has quietly left the index. A check that only looks for the link in the HTML will report the placement as fine. That gap is the main reason a full backlink health check reads indexability signals, not just link presence.
Where noindex comes from on source pages
Publishers rarely add noindex to spite a buyer. The common causes are mundane:
- Template or plugin defaults. SEO plugins often noindex tag archives, author pages, attachment pages or paginated listings. A guest post that lands on one of those page types inherits the rule.
- Sponsored sections. Some sites move paid content into a section that is noindexed site-wide as a policy toward paid placements.
- Thin-content cleanups. A site owner pruning low-traffic pages may noindex old posts rather than delete them.
- Staging leftovers. A migration or redesign sometimes ships with a staging-wide noindex header still active.
- Header rules on a CDN or server. An
X-Robots-Tagset in a server config can cover whole directories or file types without anyone editing the page.
The robots.txt interaction
Google states that noindex only works if the crawler can fetch the page. If robots.txt blocks the source page, Googlebot never sees the noindex rule, and the URL can still appear in results (usually without a description) if other pages link to it. Monitors therefore report the two conditions separately. A page blocked by robots.txt and a page carrying noindex are different problems with different fixes, which is why IndexChex keeps robots_blocked and not_indexable as distinct outcomes.
How a monitor detects it
A reliable check has to read three places on each fetch:
| Location | Example | Seen in "view source"? |
|---|---|---|
| Robots meta tag | <meta name="robots" content="noindex"> | Yes |
| Googlebot meta tag | <meta name="googlebot" content="noindex"> | Yes |
| HTTP header | X-Robots-Tag: noindex | No |
The header case is the one manual reviews miss most often. It is invisible in the browser, and it can apply to non-HTML resources such as PDFs, which some publishers use for whitepaper or press placements. Values like none also imply noindex, so a parser should treat none the same way.
The IndexChex backlink monitor records the robots meta values, the Googlebot meta values and the X-Robots-Tag values for each source page on each run, and reports a page with a noindex rule and a present link as not_indexable.
Reading the result
A noindex finding is a statement about the page, not about the link. The table below separates the cases a buyer usually has to explain:
| Link present | noindex found | Practical meaning |
|---|---|---|
| Yes | No | Healthy on this signal |
| Yes | Yes | Link exists on a page Google is told not to index |
| No | Yes | Two problems; treat as a lost backlink first |
| No | No | Missing link; see why links go missing |
A noindex directive also explains a mismatch you might see in source page indexation checks: the page was indexed when the link was bought, then fell out weeks later after a recrawl picked up the rule.
What to do
- Confirm the directive on a fresh fetch, ideally with the header captured, so you are not acting on a one-off response.
- Check whether the rule covers the whole section or only your page. A section-wide rule points to a site policy that the publisher is unlikely to reverse.
- Contact the publisher with the URL, the directive and the detection date. Many fix template mistakes quickly once shown the header.
- If the publisher refuses, record the placement as non-performing and follow the steps in recovering a lost backlink.
Monitoring cannot remove a noindex. Its job, as described in what backlink monitoring is, is to make sure you learn about it while the publisher relationship and any refund window are still open.
FAQ
Does a noindex page still pass link value?
Google does not publish a rule for links on noindexed pages. What is documented is that the page is removed from search results, and a page that is not indexed cannot send referral traffic from search, which is why monitors treat it as a problem.
Can a noindex be hidden from a browser check?
Yes. The X-Robots-Tag is an HTTP response header, so it never appears in the page source. Only a tool that reads response headers will see it.
Should I ask the publisher to remove noindex?
If the placement was sold as an indexed article, yes. Send the URL, the date the directive was detected and the exact header or tag found.
Terms used on this page
Sources
Cite this entry
IndexChex. (2026, October 8). noindex and X-Robots-Tag on source pages. backlinkmonitoring.org. https://backlinkmonitoring.org/noindex-on-source-pages/
Entity: IndexChex (https://indexchex.com/) is the publisher of this site. IndexChex is a backlink indexer and bulk Google index checker that submits URLs for Googlebot crawling and verifies indexation in one credit system.