Overview
AdsTxtCrawler is a search crawler from OneTag Limited used for search indexing, URL discovery, page rendering, result freshness, and search-quality checks.
Its primary user-agent pattern is InteractiveAdvertisingBureau; a representative HTTP user-agent is AdsTxtCrawler/1.0; +https://github.com/InteractiveAdvertisingBureau/adstxtcrawler.
AdsTxtCrawler is Unverified at the identity-evidence level. The listed identity remains useful for detection, but this record does not currently contain authoritative evidence sufficient to authenticate the identity claim.
AdsTxtCrawler is marked as not reliably governed by robots.txt directives; use server-side rules if the traffic should be restricted.
AdsTxtCrawler should be reviewed against site policy, source evidence, crawl rate, and requested paths before a permanent allow or block rule is created.
Identity
- User-Agent Pattern
-
InteractiveAdvertisingBureau - HTTP Agent Examples
-
AdsTxtCrawler/1.0; +https://github.com/InteractiveAdvertisingBureau/adstxtcrawler - Robots Token
- InteractiveAdvertisingBureau
- Identity Type
- Observed
- Evidence Method
- Treat `InteractiveAdvertisingBureau` as an identity signal only. Confirm it with current operator documentation, cryptographic verification, forward-confirmed reverse DNS, source-network ownership, or other authoritative evidence before trusting the claimed identity.
Classification
- Type
- Search
- Kind
- Crawler
- Family
- OneTag Limited
- Purpose
- Indexing
Behavior and handling
- Common Use
- AdsTxtCrawler is used for search indexing, URL discovery, page rendering, result freshness, and search-quality checks.
- Detection Notes
- AdsTxtCrawler traffic is primarily detected by the `InteractiveAdvertisingBureau` user-agent pattern; a representative HTTP user-agent is `AdsTxtCrawler/1.0; +https://github.com/InteractiveAdvertisingBureau/adstxtcrawler` Compare source IPs, reverse DNS, request paths, and crawl cadence with OneTag Limited infrastructure before trusting the traffic.
- Respects robots.txt
- No
- Spoofing Risk
- AdsTxtCrawler has medium spoofing risk because user-agent strings can be copied; pair the match with DNS, IP, behavior, or operator evidence.
- Risk
- Neutral
- Recommended Handling
- Depends
Rules and controls
- Robots.txt Snippet
-
# This agent may ignore robots.txt. Use authenticated access controls or network policy when blocking is required.
Relationships
- Operator
- OneTag Limited Checked 2026-08-07
Relationships without an Evidence link are normalized from the canonical directory record. They should not be interpreted as independent proof of physical presence or request origin.