What is BeansLLM-CorpusBot?
BeansLLM-CorpusBot is a generic crawler. Use this crawler profile to identify user-agent tokens, operator signals, platform hints, and recommended handling.
BeansLLM-CorpusBot
Generic crawler
- Version
- 1.1
- First seen
- Oct 7, 2026
- Confidence
- Inferred user-agent token
Crawler tags
What is BeansLLM-CorpusBot?
BeansLLM-CorpusBot is a generic crawler. BeansLLM-CorpusBot is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.
BeansLLM-CorpusBot was inferred from a bot-like product token in the user-agent string.
How to identify BeansLLM-CorpusBot in logs
Search server logs for BeansLLM-CorpusBot, robots.txt. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.
BeansLLM-CorpusBot/1.1 (+https://scrape.beansllm.ai/bot; respects robots.txt)
Use this as a web-crawler lookup reference for identifying how this user-agent presents itself in server logs. After you identify crawler traffic, run the AI Agent Readiness Scanner to confirm whether AI crawlers and agents can understand your site.
User-agent signals
- Product tokens
- BeansLLM-CorpusBot/1.1, respects, robots.txt
- Documentation
- https://scrape.beansllm.ai/bot
- Contact
- None found
- Platform
- Unknown · Unknown
- Browser profile
- Unknown · Unknown
- Browser-like UA
- No
- HTTP library
- Unknown
- Spoof risk
- High
Questions answered by this crawler profile
What is BeansLLM-CorpusBot?
BeansLLM-CorpusBot is classified as a generic crawler that may fetch pages for search, SEO, monitoring, or data collection workflows.
Who operates BeansLLM-CorpusBot?
The operator for BeansLLM-CorpusBot is not known from the user-agent alone.
How do I identify BeansLLM-CorpusBot in logs?
Search server logs for BeansLLM-CorpusBot, robots.txt. Matching those tokens is useful for discovery, but IP verification is still recommended before trusting the identity.
Should I allow BeansLLM-CorpusBot?
Verify IP ownership or behavior before making security decisions because user-agent strings can be spoofed. Monitor crawl rate and paths, then allow normal traffic or rate-limit/block if behavior becomes abusive.
Does BeansLLM-CorpusBot respect robots.txt?
Robots.txt compliance cannot be proven from a user-agent string alone. Check the crawler operator documentation and your own logs before assuming behavior.
