이 MCP로 할 수 있는 일
Analyzes public webpages for robots rules, bot walls, machine-readable content, sources, licenses, and embedded structured data.
도구
입력 스키마
{'type': 'object', 'title': 'can_fetch_toolArguments', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'items': {'type': 'string'}, 'title': 'Urls', 'description': '1 to 100 public http(s) URLs to check in one payment. Robots.txt is fetched once per distinct origin, not once per URL (at most 25 distinct origins per call). An unreadable or unparsable robots.txt is reported as unknown, never as allowed.'}, 'agent': {'type': 'string', 'title': 'Agent', 'default': '*', 'description': "Crawler user-agent token to evaluate against robots.txt, e.g. 'GPTBot'. Case-insensitive, stripped and truncated to 64 characters. Optional: omit it, or pass an empty string, to evaluate the wildcard ('*') group."}}}
입력 스키마
{'type': 'object', 'title': 'crawl_policy_toolArguments', 'properties': {'limit': {'anyOf': [{'type': 'integer', 'maximum': 24, 'minimum': 1}, {'type': 'null'}], 'title': 'Limit', 'default': None, 'description': 'How many origins of the fixed seed to include, 1-24 (default: all of them). The seed list is versioned and published with the dataset.'}}}
입력 스키마
{'type': 'object', 'title': 'domain_report_toolArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'page_dataArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'page_postureArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'page_report_toolArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'page_sourcesArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'pages_posture_toolArguments', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'items': {'type': 'string'}, 'title': 'Urls', 'description': '1 to 10 public http(s) URLs to analyse in one payment. Each is analysed independently: an unreachable or private URL never fails the batch, it is reported as its own error entry in results. Private, loopback and link-local addresses are refused per URL.'}}}
입력 스키마
{'type': 'object', 'title': 'posture_previewArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}}}
입력 스키마
{'type': 'object', 'title': 'robots_check_toolArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url', 'description': 'Public http(s) URL to analyse. Must be reachable from the internet; private, loopback and link-local addresses are refused (target_not_analysable) and are never charged for.'}, 'agent': {'type': 'string', 'title': 'Agent', 'default': '*', 'description': "Crawler user-agent token to evaluate against robots.txt, e.g. 'GPTBot'. Case-insensitive, stripped and truncated to 64 characters. Optional: omit it, or pass an empty string, to evaluate the wildcard ('*') group."}}}
최근 도구 변경
유사한 MCP 서버
apis-io
Provides catalog search, comparison, scoring, enrichment, lists, and dataset exports for APIs, providers, specifications, workflo…
Tools for Agents
Provides web research and extraction workflows, crawling, browser automation, PDF and book OCR, text processing, embeddings, and …
canton-ccpedia
Searches and explains Canton Network and Daml documentation, governance proposals, forums, code examples, APIs, releases, and eco…
webbersites-x402-data-api
Provides pay-per-call utilities for web scraping, SEO and accessibility audits, document extraction, DNS and email checks, crypto…
api-evangelist
Provides API governance analysis, OpenAPI overlays and version diffs, security-sensitive field classification, code samples, matu…
arch-tools-mcp
Provides general-purpose agent APIs for browser automation, web and PDF extraction, data conversion, email operations, fact check…
Agent Margin Router
Provides pay-per-call tools for blockchain balances, contracts, prices, gas, domains, web extraction, package intelligence, news,…
1cent Web Intelligence for AI Agents
Performs SSRF-safe public web inspection, URL extraction, metadata and content analysis, accessibility checks, feed and sitemap d…