Overview
GeistHaus PageFetcher is an AI crawler from GeistHaus used for assistant-driven browsing, page retrieval, and automated web access on behalf of an AI product.
Its primary user-agent pattern is GeistHaus-PageFetcher.
GeistHaus PageFetcher is not independently verified with Low confidence. The identity type is Observed, and the evidence basis is observed traffic patterns and user-agent evidence.
Robots.txt behavior is not currently confirmed.
GeistHaus PageFetcher should be handled according to the site owner’s AI crawler policy, with allow, block, or rate-limit rules applied deliberately.
Identity
- User-Agent
GeistHaus-PageFetcher- Robots.txt Token
GeistHaus-PageFetcher- Identity Type
- Observed
- Evidence Method
- Verify GeistHaus PageFetcher by matching `GeistHaus-PageFetcher` to GeistHaus evidence, then checking reverse DNS, source-network ownership, signed request data, or published crawler documentation when available.
Classification
- Type
- AI
- Kind
- Fetcher
- Family
- GeistHaus
- Purpose
- ai-assistant
Behavior and handling
- Common Use
- GeistHaus PageFetcher is used for assistant-driven browsing, page retrieval, and automated web access on behalf of an AI product.
- Detection Notes
- GeistHaus PageFetcher traffic is primarily detected by the `GeistHaus-PageFetcher` user-agent pattern. Compare source IPs, reverse DNS, request paths, and crawl cadence with GeistHaus infrastructure before trusting the traffic.
- Respects robots.txt
- Unknown
- Spoofing Risk
- GeistHaus PageFetcher has high spoofing risk because the pattern is low-confidence or observation-based; do not trust the user-agent by itself.
- Risk
- Neutral
- Recommended Handling
- Depends
Rules and controls
- Robots.txt Snippet
# robots.txt behavior is unconfirmed. Do not rely on this rule without verification.