What is Noindex Tag?
A directive (meta robots or X-Robots-Tag) asking engines not to index a page. Correct for private, duplicate, staging, and utility pages; catastrophic when leaked onto money pages.
Related terms
Robots.txt
A plain-text file at the domain root telling compliant crawlers which paths they may fetch. It controls crawling, not indexing: blocked URLs can still be indexed by URL if linked.
Canonical Tag
A hint (rel="canonical") naming the preferred URL when duplicates exist. Only about three in five pages web-wide carry a valid one, so validation must cover target quality and conflicting signals.
Staging Site Indexation
Preview or staging hosts accidentally indexed and competing with production. Prevention is authentication plus noindex on non-production hosts; cleanup needs removal requests after the leak is sealed.
Technical SEO
The crawling, indexing, rendering, and performance work that lets search engines reach, understand, and serve your pages. It covers directives, architecture, speed, and structured data rather than words on the page.
Crawling
How search-engine bots discover pages by following links, sitemaps, and directives. If a URL is never crawled, it cannot be indexed no matter how good its content is.
Crawl Budget
The number of pages a search engine will crawl on a site in a given period, driven by server capacity and demand signals. Large or slow sites must spend it on indexable, valuable URLs.