Serveur MCP

AgentScrape

io.github.hshintelligence/agent-scrape
Outils développeur Recherche et exploration Public et accessible MCP 2025-11-25

Ce que fait ce MCP

Scrapes webpages, extracts metadata and structured data, runs browser workflows, and captures webpage screenshots.

create_browser_session
Create a stateful browser session that persists cookies and localStorage across multiple scrape/workflow calls.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'properties': {'ttl_seconds': {'type': 'number', 'description': 'Session TTL (default 1800, max 7200)'}}}
extract_metadata
Extract page metadata: title, description, Open Graph, Twitter cards, JSON-LD, canonical URL, and all meta tags.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['url'], 'properties': {'url': {'type': 'string'}}}
extract_structured_data
AI-powered structured data extraction from any webpage using natural language. Returns JSON matching your prompt or schema.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['url', 'prompt'], 'properties': {'url': {'type': 'string', 'description': 'The URL to extract from'}, 'prompt': {'type': 'string', 'description': 'Natural language description of what to extract'}, 'schema': {'type': 'object', 'description': 'Optional JSON schema for the response', 'propertyNames': {'type': 'string'}, 'additionalProperties': {}}, 'wait_ms': {'type': 'number'}, 'wait_for': {'type': 'string'}}}
run_workflow
Execute a multi-step browser workflow atomically: navigate, click, type, wait, scroll, screenshot, extract, evaluate. Up to 20 steps.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['steps'], 'properties': {'steps': {'type': 'array', 'items': {'type': 'object', 'required': ['action'], 'properties': {'ms': {'type': 'number'}, 'url': {'type': 'string'}, 'text': {'type': 'string'}, 'action': {'enum': ['navigate', 'click', 'type', 'wait_for', 'wait_ms', 'scroll', 'screenshot', 'extract', 'extract_ai', 'evaluate'], 'type': 'string'}, 'format': {'enum': ['markdown', 'html', 'text'], 'type': 'string'}, 'prompt': {'type': 'string'}, 'script': {'type': 'string'}, 'selector': {'type': 'string'}, 'full_page': {'type': 'boolean'}}}, 'description': 'Ordered list of workflow steps to execute'}, 'viewport': {'enum': ['desktop', 'mobile', 'tablet'], 'type': 'string'}, 'session_id': {'type': 'string'}, 'persist_session': {'type': 'boolean'}}}
scrape_webpage
Scrape any webpage and return content as markdown, html, text, or json. Pay-per-call web scraping for AI agents.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['url'], 'properties': {'url': {'type': 'string', 'description': 'The URL to scrape (http or https)'}, 'format': {'enum': ['markdown', 'html', 'text', 'json'], 'type': 'string', 'description': 'Output format (default: markdown)'}, 'wait_ms': {'type': 'number', 'description': 'Milliseconds to wait after page load (max 10000)'}, 'viewport': {'enum': ['desktop', 'mobile', 'tablet'], 'type': 'string', 'description': 'Viewport size (default: desktop)'}, 'wait_for': {'type': 'string', 'description': 'CSS selector to wait for before extracting'}}}
screenshot_webpage
Capture a PNG screenshot of any webpage. Supports desktop, mobile, and tablet viewports, plus full-page mode.
Schéma d’entrée
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['url'], 'properties': {'url': {'type': 'string'}, 'wait_ms': {'type': 'number'}, 'viewport': {'enum': ['desktop', 'mobile', 'tablet'], 'type': 'string'}, 'wait_for': {'type': 'string'}, 'full_page': {'type': 'boolean', 'description': 'Capture full scrollable page (default: false)'}}}
Ajouté
run_workflow
17 September 2026 12:42
Ajouté
create_browser_session
17 September 2026 12:42
Ajouté
extract_metadata
17 September 2026 12:42
Ajouté
screenshot_webpage
17 September 2026 12:42
Ajouté
extract_structured_data
17 September 2026 12:42
Ajouté
scrape_webpage
17 September 2026 12:42