Serveur MCP

Stringer CleanExtract

app.getstringer/cleanextract
Outils développeur Public et accessible MCP 2026-07-28

Ce que fait ce MCP

Extracts clean, token-dense Markdown from public URLs or raw HTML.

clean_extract
Extract a full page or selected sections from a public URL or raw HTML. A call is charged only when it returns usable content; the first 3 usable extraction calls are free, total, then USD 0.05. Use the free clean_extract_outline first. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used. Every success reports extraction_quality and extraction_content_ratio_band.
Schéma d’entrée
{'type': 'object', 'required': ['url_or_html'], 'properties': {'sections': {'type': 'array', 'items': {'type': 'string', 'pattern': '^s[1-9][0-9]*$'}, 'maxItems': 50, 'minItems': 1, 'description': 'Optional ids from clean_extract_outline. Return only these sections in document order.'}, 'url_or_html': {'type': 'string', 'minLength': 1, 'description': 'A public HTTP(S) URL or a raw HTML string to convert into Markdown.'}, 'max_output_bytes': {'type': 'integer', 'maximum': 1048576, 'minimum': 1, 'description': 'Optional maximum UTF-8 byte length of the returned Markdown.'}}, 'additionalProperties': False}
clean_extract_outline
Free outline with section ids, headings, byte sizes, and a fingerprint, without page body text or payment. A later clean_extract call can select sections and is charged only when it returns usable content. Uncharged failures name upstream_timeout, upstream_http_error, bot_challenge, js_shell, empty_extraction, or sections_not_found. Recoverable structured data names fallback_used.
Schéma d’entrée
{'type': 'object', 'required': ['url_or_html'], 'properties': {'url_or_html': {'type': 'string', 'minLength': 1, 'description': 'A public HTTP(S) URL or a raw HTML string to convert into Markdown.'}}, 'additionalProperties': False}
Ajouté
clean_extract_outline
27 September 2026 02:51
Modifié
clean_extract
27 September 2026 02:51
Modifié
clean_extract
23 September 2026 02:51
Ajouté
clean_extract
17 September 2026 07:57