Overview
Crawlson is a search crawler used for search indexing, URL discovery, page rendering, result freshness, and search-quality checks.
Its primary user-agent pattern is Crawlson; related patterns include Crawlson/; a representative HTTP user-agent is Mozilla/5.0 (compatible; Crawlson/1.0; +https://www.crawlson.com/domain).
Crawlson is verified with Medium confidence. The identity type is Verified Bot, and the evidence basis is a public crawler reference or source-linked documentation.
Crawlson is marked as respecting robots.txt directives for crawler access control.
Crawlson can usually be allowed after confirming the source and monitoring request volume.
Identity
- User-Agent
Crawlson- Aliases
- Crawlson/
- HTTP Agent Examples
Mozilla/5.0 (compatible; Crawlson/1.0; +https://www.crawlson.com/domain)- Robots.txt Token
Crawlson- Identity Type
- Verified Bot
- Evidence Method
- Verify Crawlson by matching `Crawlson` to a public crawler reference or source-linked documentation, then checking reverse DNS, IP ownership, request behavior, and crawl consistency.
Classification
- Type
- Search
- Kind
- Crawler
- Family
- Crawlson
- Purpose
- indexing
Behavior and handling
- Common Use
- Crawlson is used for search indexing, URL discovery, page rendering, result freshness, and search-quality checks.
- Detection Notes
- Crawlson traffic is primarily detected by the `Crawlson` user-agent pattern; related patterns include `Crawlson/`; a representative HTTP user-agent is `Mozilla/5.0 (compatible; Crawlson/1.0; +https://www.crawlson.com/domain)`. Compare source IPs, reverse DNS, request paths, and crawl cadence before trusting the traffic.
- Respects robots.txt
- Yes
- Spoofing Risk
- Crawlson has medium spoofing risk because user-agent strings can be copied; pair the match with DNS, IP, behavior, or operator evidence.
- Risk
- Safe
- Recommended Handling
- Depends
Rules and controls
- Robots.txt Snippet
User-agent: Crawlson Disallow: /