MCP 서버

scrapewright

io.github.Ozymandias-Owens-2/scrapewright
데이터 및 분석 공개 · 연결 가능 MCP 2025-11-25

이 MCP로 할 수 있는 일

Detects website platforms and extracts structured data from individual pages or crawled sites using reusable parsing recipes.

account
Credits left and this month's usage for the key in use.
입력 스키마
{'type': 'object', 'title': 'accountArguments', 'properties': {}}
출력 스키마
{'type': 'object', 'title': 'accountDictOutput', 'additionalProperties': True}
crawl_site
Walk a site from one listing URL and extract every item. Waits up to four minutes; a longer crawl returns a job_id to pass to crawl_status.
입력 스키마
{'type': 'object', 'title': 'crawl_siteArguments', 'required': ['listing_url'], 'properties': {'js': {'type': 'boolean', 'title': 'Js', 'default': False}, 'fields': {'anyOf': [{'type': 'array', 'items': {'type': 'string'}}, {'type': 'null'}], 'title': 'Fields', 'default': None}, 'scroll': {'type': 'integer', 'title': 'Scroll', 'default': 0}, 'max_items': {'type': 'integer', 'title': 'Max Items', 'default': 25}, 'listing_url': {'type': 'string', 'title': 'Listing Url'}}}
출력 스키마
{'type': 'object', 'title': 'crawl_siteDictOutput', 'additionalProperties': True}
crawl_status
Fetch a crawl that outlived its call.
입력 스키마
{'type': 'object', 'title': 'crawl_statusArguments', 'required': ['job_id'], 'properties': {'job_id': {'type': 'string', 'title': 'Job Id'}}}
출력 스키마
{'type': 'object', 'title': 'crawl_statusDictOutput', 'additionalProperties': True}
detect_site
Report what platform a site runs on and which strategy to use. Cheap; call it before a large job.
입력 스키마
{'type': 'object', 'title': 'detect_siteArguments', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url'}}}
출력 스키마
{'type': 'object', 'title': 'detect_siteDictOutput', 'additionalProperties': True}
extract_page
Extract structured data from ONE page. ``fields`` declares your own schema, e.g. ["title", "salary:number", "tags:list"]; omit it for the product schema. First call on a new site compiles a recipe (300 credits); later calls replay it for 1 credit per row.
입력 스키마
{'type': 'object', 'title': 'extract_pageArguments', 'required': ['url'], 'properties': {'js': {'type': 'boolean', 'title': 'Js', 'default': False}, 'url': {'type': 'string', 'title': 'Url'}, 'fields': {'anyOf': [{'type': 'array', 'items': {'type': 'string'}}, {'type': 'null'}], 'title': 'Fields', 'default': None}}}
출력 스키마
{'type': 'object', 'title': 'extract_pageDictOutput', 'additionalProperties': True}
추가됨
account
2026년 9월 21일 2:40 AM
추가됨
crawl_status
2026년 9월 21일 2:40 AM
추가됨
crawl_site
2026년 9월 21일 2:40 AM
추가됨
extract_page
2026년 9월 21일 2:40 AM
추가됨
detect_site
2026년 9월 21일 2:40 AM