MCP Server

mcp

com.scrapingant/mcp
Cloud & Infrastructure Developer Tools Public & reachable MCP 2025-11-25

What this MCP does

Fetches web pages through cloud-hosted browsers with JavaScript rendering, proxies, and HTML, Markdown, or text output.

get_web_page_html
Fetch (scrape) a URL using ScrapingAnt and return the web page content as HTML. Args: url: The URL of the page to extract (scrape). browser: Whether to use browser rendering. Default: True. proxy_type: Type of proxy to use. Default: 'datacenter'. Use 'residential' if you encounter anti-bot detection, which improves anti-bot avoidance. proxy_country: Optional ISO-3166 country code. Default: random worldwide proxy. Use when facing geo-restrictions. Available country codes: ae, br, bz, ca, cn, cz, de, es, fr, gb, hk, id, il, in, it, jp, kr, mx, my, nh, nl, ph, pk, pl, ro, ru, sa, sc, se, sg, th, tr, tw, uk, us, vn.
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url'}, 'browser': {'type': 'boolean', 'title': 'Browser', 'default': True}, 'proxy_type': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Type', 'default': None}, 'proxy_country': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Country', 'default': None}}}
Output schema
{'type': 'object', 'title': '_WrappedResult', 'required': ['result'], 'properties': {'result': {'type': 'string', 'title': 'Result'}}, 'description': 'Generic wrapper for non-object return types.', 'x-fastmcp-wrap-result': True}
get_web_page_markdown
Fetch (scrape) a URL using ScrapingAnt and return the web page content as Markdown. Args: url: The URL of the page to extract (scrape). browser: Whether to use browser rendering. Default: True. proxy_type: Type of proxy to use. Default: 'datacenter'. Use 'residential' if you encounter anti-bot detection, which improves anti-bot avoidance. proxy_country: Optional ISO-3166 country code. Default: random worldwide proxy. Use when facing geo-restrictions. Available country codes: ae, br, bz, ca, cn, cz, de, es, fr, gb, hk, id, il, in, it, jp, kr, mx, my, nh, nl, ph, pk, pl, ro, ru, sa, sc, se, sg, th, tr, tw, uk, us, vn.
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url'}, 'browser': {'type': 'boolean', 'title': 'Browser', 'default': True}, 'proxy_type': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Type', 'default': None}, 'proxy_country': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Country', 'default': None}}}
Output schema
{'type': 'object', 'title': '_WrappedResult', 'required': ['result'], 'properties': {'result': {'type': 'string', 'title': 'Result'}}, 'description': 'Generic wrapper for non-object return types.', 'x-fastmcp-wrap-result': True}
get_web_page_text
Fetch (scrape) a URL using ScrapingAnt and return the web page content as plain text. Args: url: The URL of the page to extract (scrape). browser: Whether to use browser rendering. Default: True. proxy_type: Type of proxy to use. Default: 'datacenter'. Use 'residential' if you encounter anti-bot detection, which improves anti-bot avoidance. proxy_country: Optional ISO-3166 country code. Default: random worldwide proxy. Use when facing geo-restrictions. Available country codes: ae, br, bz, ca, cn, cz, de, es, fr, gb, hk, id, il, in, it, jp, kr, mx, my, nh, nl, ph, pk, pl, ro, ru, sa, sc, se, sg, th, tr, tw, uk, us, vn.
Input schema
{'type': 'object', 'required': ['url'], 'properties': {'url': {'type': 'string', 'title': 'Url'}, 'browser': {'type': 'boolean', 'title': 'Browser', 'default': True}, 'proxy_type': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Type', 'default': None}, 'proxy_country': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Proxy Country', 'default': None}}}
Output schema
{'type': 'object', 'title': '_WrappedResult', 'required': ['result'], 'properties': {'result': {'type': 'string', 'title': 'Result'}}, 'description': 'Generic wrapper for non-object return types.', 'x-fastmcp-wrap-result': True}
Added
get_web_page_text
Sept. 17, 2026, 12:37 p.m.
Added
get_web_page_markdown
Sept. 17, 2026, 12:37 p.m.
Added
get_web_page_html
Sept. 17, 2026, 12:37 p.m.