Web crawler lookup

What is Arquivo-web-crawler?

Arquivo-web-crawler is listed as a archive crawler. Use this crawler profile to identify user-agent tokens, operator signals, platform hints, and recommended handling.

Arquivo-web-crawler

Archive crawler

Version
Unknown
First seen
Sep 17, 2025
Confidence
Known user-agent token

Crawler tags

Archive crawlerNot AI trainingObeys robots.txtSystem-triggered

Directory facts

AI model training
Not listed as training
Acts on behalf of user
No, system-triggered
Obeys directives
Yes, listed as obeying robots.txt

What is Arquivo-web-crawler?

Arquivo-web-crawler is listed as a archive crawler. Arquivo-web-crawler is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

Arquivo-web-crawler matched a known crawler token in the user-agent string, but this page alone does not prove IP ownership.

How to identify Arquivo-web-crawler in logs

Search server logs for Arquivo-web-crawler. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

Arquivo-web-crawler

Arquivo-web-crawler (compatible; heritrix/3.4.0-20200304 +https://arquivo.pt/faq-crawling)

Use this as a web-crawler lookup reference for identifying how this user-agent presents itself in server logs. After you identify crawler traffic, run the AI Agent Readiness Scanner to confirm whether AI crawlers and agents can understand your site.

User-agent signals

Product tokens
Arquivo-web-crawler, heritrix/3.4.0-20200304
Contact
None found
Platform
Unknown · Unknown
Browser profile
Unknown · Unknown
Browser-like UA
No
HTTP library
Unknown
Spoof risk
Medium

Questions answered by this crawler profile

What is Arquivo-web-crawler?

Arquivo-web-crawler is listed as a archive crawler. Arquivo-web-crawler is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

Does Arquivo-web-crawler train AI models?

Arquivo-web-crawler is not listed as being used to train AI or LLM systems.

Is Arquivo-web-crawler user-triggered?

Arquivo-web-crawler is listed as operating independently of a direct user action.

How do I identify Arquivo-web-crawler in logs?

Search server logs for Arquivo-web-crawler. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

Does Arquivo-web-crawler respect robots.txt?

This directory entry lists the crawler as obeying robots.txt directives. Confirm behavior in your logs before relying on user-agent strings alone.