Web Intel
Ce que fait ce MCP
Offers web extraction and monitoring tools for domain health, structured page data, Markdown conversion, sitemaps, technology stacks, and storefront intelligence.
Outils
Schéma d’entrée
{'type': 'object', 'title': 'Domain Health Checker input', 'required': ['domains'], 'properties': {'domains': {'type': 'array', 'title': 'Domains', 'editor': 'stringList', 'prefill': ['example.com', 'example.com', 'github.com'], 'maxItems': 100, 'description': 'List of domains to audit (e.g. `example.com`).'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 10, 'maximum': 50, 'minimum': 1, 'description': 'How many domains to check in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'properties': {}}
Schéma d’entrée
{'type': 'object', 'title': 'Shopify Price Change Monitor input', 'required': ['websites'], 'properties': {'websites': {'type': 'array', 'title': 'Websites', 'editor': 'stringList', 'prefill': ['allbirds.com'], 'maxItems': 10, 'description': 'Shopify (or possibly-Shopify) store URLs to watch, e.g. "allbirds.com". Every scheduled run re-checks these same sites and reports ONLY what changed since a previous run of this watch: a store newly added to the watch, or an already-tracked store\'s price range / catalog size / estimated revenue band shifting.'}, 'max_items': {'type': 'integer', 'title': 'Max change rows per run', 'default': 20, 'maximum': 200, 'minimum': 1, 'description': 'Caps how many new-store / changed-store rows a single run will deliver and charge for, even if more were found.'}, 'baseline_key': {'type': 'string', 'title': 'Watch name', 'editor': 'textfield', 'prefill': 'apify-daily-test', 'description': 'A name for THIS watch, so you can run several independent store watches from one Actor (e.g. "dtc-competitors", "my-portfolio") without one overwriting another\'s memory of what\'s already been seen. Each name is scoped to YOUR OWN Apify account. The prefilled value is only there so this Actor\'s own daily test run has a stable, obviously-a-test name; replace it with your own watch name.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Shopify Store Intelligence input', 'required': ['websites'], 'properties': {'websites': {'type': 'array', 'title': 'Websites', 'editor': 'stringList', 'prefill': ['https://allbirds.com', 'https://gymshark.com', 'https://stripe.com'], 'maxItems': 100, 'description': 'List of websites to check (e.g. `allbirds.com` or `https://example.com`). One row per site.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 5, 'maximum': 20, 'minimum': 1, 'description': 'How many websites to check in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Sitemap to Knowledge input', 'required': ['items'], 'properties': {'items': {'type': 'array', 'title': 'Domains / websites', 'editor': 'stringList', 'prefill': ['docs.example.com', 'example.com'], 'maxItems': 10, 'description': 'List of domains or website URLs to crawl via their sitemap. One entry per site.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 5, 'maximum': 20, 'minimum': 1, 'description': 'How many SITES to process in parallel (each site already fetches up to 25 pages internally).'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Structured Extract input', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'title': 'URLs', 'editor': 'stringList', 'prefill': ['https://example.com', 'https://www.iana.org/help/example-domains'], 'maxItems': 100, 'description': 'List of URLs to extract structured data from. One row per URL.'}, 'fields': {'type': 'array', 'title': 'Fields to extract', 'editor': 'stringList', 'prefill': [], 'maxItems': 10, 'description': 'Optional subset of fields to return: title, description, image, siteName, canonical, jsonLd, headings, links, emails, prices. Leave empty to extract all of them.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 10, 'maximum': 30, 'minimum': 1, 'description': 'How many URLs to process in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Tech-Stack Change Detector input', 'required': ['items'], 'properties': {'items': {'type': 'array', 'title': 'Domains (with optional previous stack)', 'editor': 'stringList', 'prefill': ['shopify.com', 'wordpress.org|WordPress,jQuery'], 'maxItems': 100, 'description': 'One entry per domain. Bare form: "example.com" (just detects the current stack). Diff form: "example.com|Tech1,Tech2,Tech3" — everything after the pipe is the stack you last saw for this domain; the Actor returns what was added/removed vs. right now.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 10, 'maximum': 50, 'minimum': 1, 'description': 'How many domains to check in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Company Tech-Stack Detector input', 'required': ['websites'], 'properties': {'websites': {'type': 'array', 'title': 'Websites', 'editor': 'stringList', 'prefill': ['https://shopify.com', 'https://wordpress.org', 'https://vercel.com'], 'maxItems': 100, 'description': 'List of websites to fingerprint (e.g. `shopify.com` or `https://example.com`). One row per site.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 10, 'maximum': 50, 'minimum': 1, 'description': 'How many websites to fingerprint in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'URL to Markdown input', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'title': 'URLs', 'editor': 'stringList', 'prefill': ['https://example.com', 'https://www.iana.org/help/example-domains'], 'maxItems': 100, 'description': 'List of URLs to fetch and convert to Markdown. One row per URL.'}, 'includeLinks': {'type': 'boolean', 'title': 'Include links', 'default': True, 'description': 'Convert <a href> tags to Markdown links. Turn off to strip links and keep only their text.'}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 10, 'maximum': 30, 'minimum': 1, 'description': 'How many URLs to fetch and convert in parallel.'}}, 'schemaVersion': 1}
Schéma d’entrée
{'type': 'object', 'title': 'Zid Store Products input', 'required': ['items'], 'properties': {'items': {'type': 'array', 'title': 'Zid shop subdomains', 'editor': 'stringList', 'prefill': ['furniture', 'scent'], 'maxItems': 20, 'description': "One entry per Zid storefront — the shop's subdomain (e.g. `furniture`) or full host (`furniture.zid.store`). Find the subdomain in the store's own zid.store URL, or via its custom domain's storefront (Zid stores usually keep the *.zid.store host reachable even with a custom domain attached)."}, 'maxConcurrency': {'type': 'integer', 'title': 'Max concurrency', 'default': 5, 'maximum': 20, 'minimum': 1, 'description': 'How many shops to scan in parallel.'}}, 'schemaVersion': 1}
Modifications récentes des outils
Serveurs MCP similaires
apis-io
Provides catalog search, comparison, scoring, enrichment, lists, and dataset exports for APIs, providers, specifications, workflo…
Tools for Agents
Provides web research and extraction workflows, crawling, browser automation, PDF and book OCR, text processing, embeddings, and …
canton-ccpedia
Searches and explains Canton Network and Daml documentation, governance proposals, forums, code examples, APIs, releases, and eco…
webbersites-x402-data-api
Provides pay-per-call utilities for web scraping, SEO and accessibility audits, document extraction, DNS and email checks, crypto…
api-evangelist
Provides API governance analysis, OpenAPI overlays and version diffs, security-sensitive field classification, code samples, matu…
arch-tools-mcp
Provides general-purpose agent APIs for browser automation, web and PDF extraction, data conversion, email operations, fact check…
Agent Margin Router
Provides pay-per-call tools for blockchain balances, contracts, prices, gas, domains, web extraction, package intelligence, news,…
1cent Web Intelligence for AI Agents
Performs SSRF-safe public web inspection, URL extraction, metadata and content analysis, accessibility checks, feed and sitemap d…