Web crawler lookup

What is magpie-crawler?

magpie-crawler is listed as a data collection crawler. Use this crawler profile to identify user-agent tokens, operator signals, platform hints, and recommended handling.

magpie-crawler

Data collection crawler

Version
1.1
First seen
Sep 17, 2025
Confidence
Known user-agent token

Crawler tags

Data collection crawlerNot AI trainingObeys robots.txtSystem-triggered

Directory facts

AI model training
Not listed as training
Acts on behalf of user
No, system-triggered
Obeys directives
Yes, listed as obeying robots.txt

What is magpie-crawler?

magpie-crawler is listed as a data collection crawler. magpie-crawler is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

magpie-crawler matched a known crawler token in the user-agent string, but this page alone does not prove IP ownership.

How to identify magpie-crawler in logs

Search server logs for magpie-crawler. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

magpie-crawler

magpie-crawler/1.1 (U; Linux amd64; en-GB; +http://www.brandwatch.net)

Use this as a web-crawler lookup reference for identifying how this user-agent presents itself in server logs. After you identify crawler traffic, run the AI Agent Readiness Scanner to confirm whether AI crawlers and agents can understand your site.

User-agent signals

Product tokens
magpie-crawler/1.1, Linux, amd64, en-GB
Contact
None found
Platform
Linux · Server
Browser profile
Unknown · Unknown
Browser-like UA
No
HTTP library
Unknown
Spoof risk
Medium

Questions answered by this crawler profile

What is magpie-crawler?

magpie-crawler is listed as a data collection crawler. magpie-crawler is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

Does magpie-crawler train AI models?

magpie-crawler is not listed as being used to train AI or LLM systems.

Is magpie-crawler user-triggered?

magpie-crawler is listed as operating independently of a direct user action.

How do I identify magpie-crawler in logs?

Search server logs for magpie-crawler. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

Does magpie-crawler respect robots.txt?

This directory entry lists the crawler as obeying robots.txt directives. Confirm behavior in your logs before relying on user-agent strings alone.