Glossary · Crawlers and access
Crawler trap
An unbounded set of generated URLs — faceted filters, infinite calendars, recursive relative paths — that a crawler can follow forever. It consumes crawl allocation without exposing new content.
Crawlers and access
Who is fetching, whether they are who they claim, and what your rules actually permit.
retrieval crawler · user-triggered fetch · verified crawler · forged crawler identity · unverifiable · robots.txt · user-agent group · AI opt-out · Content-Signal · crawl budget · cloaking · challenge page at 200 · uniform refusal · nonexistent-path control · blocked render resource · off-host redirect · homepage refused · FCrDNS · Conditional request · ASN blocking · operator feed · residential proxy · Google-Extended · Google-Agent · GPTBot vs OAI-SearchBot · ClaudeBot vs Claude-User · PerplexityBot vs Perplexity-User · CCBot · Bytespider · agentic traffic · crawler classification
← FCrDNS · Conditional request →
See it in the full glossary · 579 terms across 19 areas. Scan a site to see which of these apply to it.