WebLens
What this MCP does
Fetches, crawls, searches, screenshots, compares, and extracts structured information from websites and PDFs, including research and site intelligence.
Tools
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'query': {'type': 'string', 'description': 'The question to answer'}, 'sources': {'type': 'number', 'description': 'Web sources to search, fetch, and cite (1-5, default: 3)'}}}
Input schema
{'type': 'object', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'items': {'type': 'string'}, 'description': 'URLs to fetch (2-20)'}}}
Input schema
{'type': 'object', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'items': {'type': 'string'}, 'description': 'URLs to compare (2-3)'}, 'focus': {'type': 'string', 'description': 'What to focus comparison on'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'Start URL — the crawl stays on this host'}, 'limit': {'type': 'number', 'description': 'Page budget (1-25, default: 10) — charged per requested page'}, 'exclude': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Skip URLs whose path+query contains one of these substrings'}, 'include': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Only crawl URLs whose path+query contains one of these substrings'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Per-page timeout in ms (default 10000)'}, 'maxChars': {'type': 'number', 'maximum': 50000, 'minimum': 500, 'description': 'Per-page content character cap (default 8000)'}, 'maxDepth': {'type': 'number', 'description': 'Link depth from the start URL (0-3, default: 2)'}, 'respectRobots': {'type': 'boolean', 'description': 'Honour robots.txt (default: true; disable only for sites you control)'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'depth': {'enum': ['standard', 'deep'], 'type': 'string', 'description': 'Research tier: standard = 3 sub-questions / 8 sources ($0.20); deep = 5 / 12 ($0.35). Default: standard'}, 'query': {'type': 'string', 'description': 'The research question (1-500 chars)'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'Site URL to fingerprint'}}}
Input schema
{'type': 'object', 'required': ['domain'], 'properties': {'domain': {'type': 'string', 'description': 'Domain to inspect, e.g. "stripe.com". A full URL is reduced to its hostname.'}}}
Input schema
{'type': 'object', 'required': ['url', 'schema'], 'properties': {'url': {'type': 'string', 'description': 'The URL to extract from'}, 'schema': {'type': 'object', 'description': 'JSON schema defining the data structure to extract'}, 'instructions': {'type': 'string', 'description': 'Natural language instructions to guide extraction'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'URL of the PDF to extract'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'The URL to fetch'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Timeout in ms (default 10000)'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'The URL to fetch'}, 'cache': {'type': 'boolean', 'description': 'Serve from cache when available (default true; cached responses cost 70% less)'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Request timeout in ms (default 10000)'}, 'cacheTtl': {'type': 'number', 'maximum': 86400, 'minimum': 60, 'description': 'Cache TTL in seconds (default 3600)'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'The URL to fetch'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Request timeout in ms (default 15000)'}, 'waitFor': {'type': 'string', 'description': 'CSS selector to wait for before capturing content (e.g. ".content")'}}}
Input schema
{'type': 'object', 'required': ['urls'], 'properties': {'urls': {'type': 'array', 'items': {'type': 'string'}, 'description': 'URLs to fetch (1-20)'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Per-URL timeout in ms (default 10000)'}, 'maxChars': {'type': 'number', 'maximum': 50000, 'minimum': 500, 'description': 'Per-page content character cap (default 20000)'}}}
Input schema
{'type': 'object', 'required': ['target'], 'properties': {'target': {'type': 'string', 'description': 'Company name or domain to research'}}}
Input schema
{'type': 'object', 'required': ['company'], 'properties': {'focus': {'type': 'string', 'description': 'Optional focus area'}, 'company': {'type': 'string', 'description': 'Company to analyze'}, 'maxCompetitors': {'type': 'number', 'description': 'Max competitors to include (default: 5)'}}}
Input schema
{'type': 'object', 'required': ['topic'], 'properties': {'depth': {'enum': ['quick', 'standard', 'comprehensive'], 'type': 'string', 'description': 'Research depth (default standard)'}, 'focus': {'type': 'string', 'description': 'Optional focus area'}, 'topic': {'type': 'string', 'description': 'Market or industry topic to research'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'URL to audit'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'Site URL to map'}, 'limit': {'type': 'number', 'description': 'Maximum URLs to return (1-5000, default: 1000)'}, 'exclude': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Skip URLs whose path+query contains one of these substrings'}, 'include': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Only URLs whose path+query contains one of these substrings'}, 'timeout': {'type': 'number', 'maximum': 30000, 'minimum': 5000, 'description': 'Timeout in ms (default 10000)'}}}
Input schema
{'type': 'object', 'required': ['key', 'value'], 'properties': {'key': {'type': 'string', 'description': 'Storage key (max 256 chars)'}, 'ttl': {'type': 'number', 'description': 'Time to live in hours (1-720, default: 168)'}, 'value': {'description': 'Value to store (any JSON)'}}}
Input schema
{'type': 'object', 'required': ['url', 'webhookUrl'], 'properties': {'url': {'type': 'string', 'description': 'URL to monitor for changes'}, 'webhookUrl': {'type': 'string', 'description': 'Webhook URL for change notifications'}}}
Input schema
{'type': 'object', 'required': ['name'], 'properties': {'name': {'type': 'string', 'description': 'Package name, e.g. "express" or "@scope/pkg"'}, 'registry': {'enum': ['npm', 'pypi'], 'type': 'string', 'description': 'Registry to look in (default npm)'}}}
Input schema
{'type': 'object', 'required': ['endpoint'], 'properties': {'url': {'type': 'string', 'description': 'Fetch-backed endpoints only (/fetch/basic, /contents, /map): run a real truncated preview of this URL'}, 'endpoint': {'type': 'string', 'description': 'Paid endpoint path to preview, e.g. "/answer"'}}}
Input schema
{'type': 'object', 'required': ['domain'], 'properties': {'chain': {'type': 'string', 'maxLength': 32, 'minLength': 2, 'description': 'Optional chain label, e.g. base, ethereum, solana'}, 'domain': {'type': 'string', 'maxLength': 253, 'minLength': 3, 'description': 'Project website, e.g. "example.org"'}, 'tokenAddress': {'type': 'string', 'maxLength': 80, 'minLength': 20, 'description': 'Optional contract address to cross-check against the site'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'query': {'type': 'string', 'description': 'Research topic or question'}, 'resultCount': {'type': 'number', 'description': 'Number of sources to analyze (default: 5)'}}}
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'The URL to screenshot'}, 'width': {'type': 'number', 'maximum': 3840, 'minimum': 320, 'description': 'Viewport width (default 1280)'}, 'height': {'type': 'number', 'maximum': 2160, 'minimum': 240, 'description': 'Viewport height (default 720)'}, 'fullPage': {'type': 'boolean', 'description': 'Capture full page scroll'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Max suggestions (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'Partial query to complete'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'sort': {'enum': ['relevance', 'recent'], 'type': 'string', 'description': 'Ranking (default relevance)'}, 'limit': {'type': 'number', 'maximum': 50, 'minimum': 1, 'description': 'Stories to return (default 10)'}, 'query': {'type': 'string', 'description': 'What to search Hacker News for'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'Image search query'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'News search query'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'What to search for (e.g. coffee shops)'}, 'location': {'type': 'string', 'description': 'Free-text location bias, e.g. "Austin, Texas"'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'Academic search query'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10, max: 20)'}, 'query': {'type': 'string', 'description': 'Product search query'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'query': {'type': 'string', 'description': 'Topic to get trend data for'}}}
Input schema
{'type': 'object', 'required': ['query'], 'properties': {'limit': {'type': 'number', 'description': 'Number of results (default: 10)'}, 'query': {'type': 'string', 'description': 'Search query'}, 'contentChars': {'type': 'number', 'maximum': 20000, 'minimum': 500, 'description': 'Per-page content character cap (default 8000)'}, 'contentResults': {'type': 'number', 'description': 'How many top results to fetch content for (1-10, default: 5)'}, 'includeContent': {'type': 'boolean', 'description': 'Also fetch top result pages as markdown (+$0.0015/result)'}}}
Input schema
{'type': 'object', 'required': ['url', 'query'], 'properties': {'url': {'type': 'string', 'description': 'The URL to extract from'}, 'query': {'type': 'string', 'description': 'What data to extract (natural language)'}}}
Input schema
{'type': 'object', 'required': ['videoId'], 'properties': {'lang': {'type': 'string', 'description': 'Transcript language code (default: video default)'}, 'videoId': {'type': 'string', 'description': 'YouTube video ID (e.g. dQw4w9WgXcQ) or full video URL'}}}
Recent tool changes
Similar MCP servers
agentsvc.io
Offers web retrieval, scraping, screenshots, PDF and OCR processing, validation, market data, translation, and other general-purp…
FreshContext
Evaluates supplied context for freshness and provenance and retrieves research, finance, news, repository, package, company, gove…
Industrial Platform
Runs paid tools for web and document extraction, cited research briefs, sitemap discovery, metadata analysis, and website change …
oxylabs-mcp
Searches, scrapes, crawls, maps, and browser-navigates websites and sources, including specialized Google and Amazon data extract…
LiveDataLink
Exposes public-data tools covering economics, labor, banking, transport, books, legal opinions, carriers, environmental data, and…
Stratalize Healthcare
Retired Stratalize server exposing benchmark, market intelligence, compliance, financial, healthcare, legal, real estate, and eco…
Stratalize Governance
Retired Stratalize server exposing benchmark, market intelligence, compliance, financial, healthcare, legal, real estate, and eco…
Stratalize Oracle
Retired Stratalize server exposing benchmark, market intelligence, compliance, financial, healthcare, legal, real estate, and eco…