MCP Server

Siteiz

com.siteiz/siteiz
Developer Tools Public & reachable MCP 2026-07-28

What this MCP does

Checks AI crawler permissions and visibility files, explains crawler behavior, and shows how pages appear when fetched without JavaScript.

check_ai_crawlers
Check AI crawler access
Reads a website's robots.txt and reports, for 11 AI crawlers (GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, PerplexityBot, Google-Extended, Applebot-Extended, Meta-ExternalAgent, CCBot, Bytespider), whether each may reach the given page. Separates AI search crawlers (decide if the site appears in ChatGPT, Claude and Perplexity answers) from training crawlers. Also reports llms.txt and declared sitemaps.
Read only Open world Idempotent
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'Website or page address, e.g. example.com or https://example.com/pricing'}}, 'additionalProperties': False}
Output schema
{'type': 'object', 'required': ['url', 'robotsTxt', 'agents'], 'properties': {'url': {'type': ['string', 'null']}, 'agents': {'type': 'array', 'items': {'type': 'object', 'required': ['token', 'verdict', 'operator', 'role'], 'properties': {'role': {'enum': ['search', 'user', 'training', 'control'], 'type': 'string'}, 'token': {'type': 'string'}, 'verdict': {'enum': ['allowed', 'default', 'blocked'], 'type': 'string'}, 'operator': {'type': 'string'}}}}, 'llmsTxt': {'type': ['boolean', 'null']}, 'sitemaps': {'type': 'array', 'items': {'type': 'string'}}, 'robotsTxt': {'enum': ['ok', 'missing', 'unreachable', 'invalid', 'pasted'], 'type': 'string'}, 'httpStatus': {'type': ['integer', 'null']}, 'contentSignals': {'type': 'array', 'items': {'type': 'string'}}}}
explain_ai_crawler
Explain an AI crawler
Explains what an AI crawler user agent does (operator, purpose, and what blocking it in robots.txt changes). Pass a name such as GPTBot or OAI-SearchBot, or omit it to list all crawlers Siteiz tracks.
Read only Idempotent
Input schema
{'type': 'object', 'properties': {'name': {'type': 'string', 'description': 'Crawler user agent, e.g. GPTBot. Omit to list all.'}}, 'additionalProperties': False}
Output schema
{'type': 'object', 'required': ['crawlers'], 'properties': {'crawlers': {'type': 'array', 'items': {'type': 'object', 'required': ['token', 'operator', 'role', 'does'], 'properties': {'does': {'type': 'string'}, 'role': {'type': 'string'}, 'token': {'type': 'string'}, 'product': {'type': 'string'}, 'operator': {'type': 'string'}, 'roleLabel': {'type': 'string'}}}}}}
get_ai_visibility_report
Get a published AI visibility report
Returns a published, dated Siteiz AI visibility report for a well-known company's homepage (score, grade, pillar scores, top issues). Omit the company to list all published reports.
Read only Idempotent
Input schema
{'type': 'object', 'properties': {'company': {'type': 'string', 'description': 'Company name, e.g. Notion. Omit to list all reports.'}}, 'additionalProperties': False}
Output schema
{'type': 'object', 'properties': {'url': {'type': 'string'}, 'grade': {'type': 'string'}, 'score': {'type': 'integer'}, 'report': {'type': 'string'}, 'company': {'type': 'string'}, 'pillars': {'type': 'array', 'items': {'type': 'object', 'properties': {'id': {'type': 'string'}, 'label': {'type': 'string'}, 'score': {'type': 'integer'}, 'notAssessed': {'type': 'boolean'}}}}, 'reports': {'type': 'array', 'items': {'type': 'object', 'properties': {'url': {'type': 'string'}, 'grade': {'type': 'string'}, 'score': {'type': 'integer'}, 'company': {'type': 'string'}, 'scannedAt': {'type': 'string'}}}}, 'scannedAt': {'type': 'string'}}, 'description': 'Either `reports` (a list, when no company matched) or the single report fields.'}
view_page_as_ai_crawler
View a page as an AI crawler
Fetches one page the way most AI crawlers do (plain HTTP, no JavaScript) and reports what they actually receive: title, meta description, canonical, H1, heading count, structured data types, visible word count, token reading cost, and a text preview. Use it to check whether content is visible to AI without JavaScript.
Read only Open world Idempotent
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'Website or page address, e.g. example.com or https://example.com/pricing'}}, 'additionalProperties': False}
Output schema
{'type': 'object', 'required': ['finalUrl', 'status', 'words', 'fullScan'], 'properties': {'h1': {'type': 'array', 'items': {'type': 'string'}}, 'bytes': {'type': 'integer'}, 'title': {'type': ['string', 'null']}, 'words': {'type': 'integer'}, 'status': {'type': 'integer'}, 'finalUrl': {'type': 'string'}, 'fullScan': {'type': 'string'}, 'canonical': {'type': ['string', 'null']}, 'htmlTokens': {'type': 'integer'}, 'textTokens': {'type': 'integer'}, 'headingCount': {'type': 'integer'}, 'structuredData': {'type': 'array', 'items': {'type': 'string'}}, 'metaDescription': {'type': ['string', 'null']}}}
Added
get_ai_visibility_report
Oct. 5, 2026, 2:40 a.m.
Added
explain_ai_crawler
Oct. 5, 2026, 2:40 a.m.
Added
view_page_as_ai_crawler
Oct. 5, 2026, 2:40 a.m.
Added
check_ai_crawlers
Oct. 5, 2026, 2:40 a.m.