Web crawler lookup

What is BeansLLM-CorpusBot?

BeansLLM-CorpusBot is a generic crawler. Use this crawler profile to identify user-agent tokens, operator signals, platform hints, and recommended handling.

BeansLLM-CorpusBot

Generic crawler

Version
1.1
First seen
Oct 7, 2026
Confidence
Inferred user-agent token

Crawler tags

Generic crawler

What is BeansLLM-CorpusBot?

BeansLLM-CorpusBot is a generic crawler. BeansLLM-CorpusBot is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

BeansLLM-CorpusBot was inferred from a bot-like product token in the user-agent string.

How to identify BeansLLM-CorpusBot in logs

Search server logs for BeansLLM-CorpusBot, robots.txt. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

BeansLLM-CorpusBotrobots.txt

BeansLLM-CorpusBot/1.1 (+https://scrape.beansllm.ai/bot; respects robots.txt)

Use this as a web-crawler lookup reference for identifying how this user-agent presents itself in server logs. After you identify crawler traffic, run the AI Agent Readiness Scanner to confirm whether AI crawlers and agents can understand your site.

User-agent signals

Product tokens
BeansLLM-CorpusBot/1.1, respects, robots.txt
Contact
None found
Platform
Unknown · Unknown
Browser profile
Unknown · Unknown
Browser-like UA
No
HTTP library
Unknown
Spoof risk
High

Questions answered by this crawler profile

What is BeansLLM-CorpusBot?

BeansLLM-CorpusBot is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.

Who operates BeansLLM-CorpusBot?

The operator for BeansLLM-CorpusBot is not known from the user-agent alone.

How do I identify BeansLLM-CorpusBot in logs?

Search server logs for BeansLLM-CorpusBot, robots.txt. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.

Should I allow BeansLLM-CorpusBot?

Verify IP ownership or behavior before making security decisions because user-agent strings can be spoofed. Monitor crawl rate and paths, then allow normal traffic or rate-limit/block if behavior becomes abusive.

Does BeansLLM-CorpusBot respect robots.txt?

Robots.txt compliance cannot be proven from a user-agent string alone. Check the crawler operator documentation and your own logs before assuming behavior.