Overview
A bot detection API for crawlers and scrapers. Bot Detector uses a comprehensive database of known bot and crawler user agent patterns to identify automated traffic, categorizing detected bots into types like search engines, social media crawlers, SEO tools, and monitoring services. It also applies list-free heuristics to flag automated clients the database misses — missing user agents and non-browser clients such as curl and scripts — and returns a composite risk score.
Live Test Bot Detection Grounding Data Source →
The tool
Once your client is connected to the VerveContext server, this appears in its tool list as BotDetectionGroundingData. It is read-only and open-world — it fetches and never mutates anything on your side — so most clients call it without asking you to confirm.
{
"name": "BotDetectionGroundingData",
"arguments": {
"ua": "Googlebot/2.1 (+http://www.google.com/bot.html)"
}
}You do not name the tool yourself; the model picks it. Asking about Googlebot/2.1 (+http://www.google.com/bot.html) in the terms this source covers is enough for it to reach for BotDetectionGroundingData on its own — naming it explicitly also works, and is the way to force the call.
Connecting
One server URL covers every source in the catalog, including this one. Authorization is OAuth: the client opens a browser once, and there is no key to paste into a config file.
{
"mcpServers": {
"vervecontext": {
"url": "https://api.vervecontext.com/v1/mcp"
}
}
}https://api.vervecontext.com/v1/mcpPer-client setup — Claude, Cursor, VS Code, ChatGPT — is on the MCP setup page.
Arguments
These are the properties on the tool's inputSchema, so a well-behaved client validates them before the call is made. Premium arguments are accepted on every plan but only take effect on plans that include them.
| Argument | Type | Description |
|---|---|---|
uaRequired | string | The user agent string to analyze (URL encoded) |
ipOptional | string | Optional IP the request came from. When the user agent claims to be a Google, Bing, OpenAI or Apple crawler, it is checked against that operator's published IP ranges ip |
What the model gets back
The result carries a structuredContent object matching the tool's declared outputSchema, so a client reads fields without parsing prose. status is "ok" and error is null on success; a null field means the value was not available for that input, not that the call failed.
{
"status": "ok",
"error": null,
"data": {
"userAgent": "Googlebot/2.1 (+http://www.google.com/bot.html)",
"isBot": true,
"bot": {
"name": "Googlebot",
"category": "search_engine",
"url": "http://www.google.com/bot.html",
"reputation": "trusted",
"shouldBlock": false
},
"isAutomated": true,
"verified": true,
"riskScore": 0,
"riskLevel": "low"
}
}
Response fields
Paths are relative to data. Premium fields are absent rather than zeroed on plans that do not include them, so check for presence instead of comparing to 0.
| Field | Type | Example | Description |
|---|---|---|---|
userAgent | string | Googlebot/2.1 (+http://www.google.com/bot.html) | The analyzed user agent string provided in the request |
isBot | boolean | true | Whether the user agent belongs to a known bot or crawler |
botPremium | object | {…} | Detailed bot information including name, category, and reputation |
bot.namePremium | string | Googlebot | Name of the detected bot or crawler |
bot.categoryPremium | string | search_engine | Category of bot: search_engine, social_media, monitoring, seo_tool, scraper, other |
bot.urlPremium | string | http://www.google.com/bot.html | Official URL or documentation link for the bot |
bot.reputationPremium | string | trusted | Trust level: trusted (search engines, social), neutral (monitoring, SEO), malicious (scrapers, spam), or unknown |
bot.shouldBlockPremium | boolean | false | Recommendation whether to block this bot based on category and reputation |
isAutomatedPremium | boolean | true | Heuristic signal that the traffic is automated — true for known bots, missing user agents, and clients that do not present as a standard browser (e.g. curl, scripts, headless tooling). Catches automated clients not in the bot database |
verifiedPremium | boolean | true | Whether a claimed Google, Bing, OpenAI or Apple crawler really comes from that operator's published IP ranges (false = spoofed). Null when no ip is given or the user agent claims no verifiable crawler |
riskScorePremium | number | 0 | Composite 0-100 risk score combining bot reputation, block recommendation and automation heuristics (higher is riskier) |
riskLevelPremium | string | low | Risk band derived from the score: low, medium or high |
Why ground on it
A model can produce something that looks like this answer from its training data, and be confidently out of date or simply wrong. This source returns the current value in a shape you can check, which is the difference between an answer you can cite and one you have to hedge.
Point an evaluation at userAgent: it is the field most worth pinning a claim to, and it is either present and current or absent — never plausibly invented.
Failure modes
Errors come back as tool errors carrying a sentence the model can act on, not a bare status code. Error handling covers the full list.
| Status | What it means |
|---|---|
400 / 422 | The arguments did not validate. The message names the offending one. |
401 | The OAuth session is invalid or expired — reconnect the server. |
403 | Blocked by a key restriction or an IP allow-list. Never a bad identity. |
404 | This source is not part of VerveContext. Check the catalog. |
429 | Out of credits, or a brief rate limit. The message tells them apart. |
A call costs 10 credits each time the tool actually runs; a model that reasons about the tool without calling it costs nothing.
Use cases
- Analytics Filtering
- Exclude bot traffic from your analytics to get accurate visitor metrics and user behavior data
- Access Control
- Implement different behavior for bots vs real users, such as rate limiting or content restrictions
- Security Monitoring
- Identify potentially malicious crawlers and scrapers attempting to access your content
- SEO Insights
- Track which search engine bots are crawling your site and how frequently
Other ways to use Bot Detection Grounding Data
Set up Bot Detection Grounding Data on VerveContext, or reach the same source a different way. Your VerveContext account and credits work on all of them — one key, one balance.
Related
More in Networking: