MCP-Server

Brainiall Image

com.brainiall/image
Design & Kreativ Entwicklertools Medien & Inhalte Öffentlich und erreichbar MCP 2025-11-25

Was dieses MCP kann

Processes images and documents through background removal, upscaling, face restoration, OCR, table extraction, document conversion, and structured visual understanding.

check_image_service
Check Image Service
Check health status of Image API services and loaded models. Returns: dict with keys: - status (str): 'healthy' or error state - models (dict): Loaded model status per capability - version (str): API version
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'properties': {}}
document_extract
Extract Document Fields
Turn a document image into structured fields. doc_type picks the schema (receipt/invoice/id/contract/form/generic). A page with no readable text returns an error rather than a guess. Returns: dict with keys: doc_type (str), fields (dict — null for any value not present), text (str — the recognised plain text).
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image'], 'properties': {'image': {'type': 'string', 'description': 'Base64-encoded PNG/JPEG of a single document page'}, 'doc_type': {'type': 'string', 'default': 'generic', 'description': 'The document kind â\x80\x94 picks the field schema: receipt | invoice | id | contract | form | generic | business_card | w2 | health_card | mortgage | pay_stub'}}}
document_query
Ask Question About Document
Ask a natural-language question about a document image; returns a grounded answer plus the supporting line. Returns found:false rather than guessing when the document doesn't contain the answer. Returns: dict with keys: answer (str|null), found (bool), supporting_text (str|null), text (str).
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image', 'question'], 'properties': {'image': {'type': 'string', 'description': 'Base64-encoded PNG/JPEG of the document page'}, 'question': {'type': 'string', 'maxLength': 1000, 'description': 'The natural-language question about the document'}}}
document_tables
Extract Tables From Document
Reconstruct every table in a document image into headers and rows. Returns: dict with keys: table_count (int), tables (list of {title, headers, rows, row_count, column_count}); [] if there are no tables.
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image'], 'properties': {'image': {'type': 'string', 'description': 'Base64-encoded PNG/JPEG of the document page'}}}
document_to_markdown
Document to Markdown (Layout)
Return the document as structured Markdown (headings, tables, lists, code blocks, math). Brainiall Doc Layout engine. The single API for converting documents to LLM-friendly format.
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['pdf'], 'properties': {'pdf': {'type': 'string', 'description': 'Base64-encoded PDF document'}, 'page_range': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'default': None, 'description': "Page range like '1,2,5-10' or null for all pages"}}}
remove_background
Remove Background
Remove the background from an image. Uses Brainiall Cutout engine segmentation to precisely separate foreground from background. Returns a base64-encoded image with transparent background (PNG) or white background (WebP). Sub-500ms latency on GPU. Args: image_base64: Base64-encoded image data (PNG, JPEG, or WebP). output_format: Output format -- 'png' (with transparency) or 'webp'. Returns: dict with keys: - image_base64 (str): Base64-encoded result image - format (str): Output image format - original_size (dict): Original width and height - processing_ms (int): Processing time in milliseconds
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image_base64'], 'properties': {'image_base64': {'type': 'string', 'maxLength': 20000000, 'description': 'Base64-encoded image data. Supports PNG, JPEG, and WebP formats.'}, 'output_format': {'type': 'string', 'default': 'png', 'description': "Output image format: 'png' (default, with transparency) or 'webp'"}}}
restore_face
Restore Face
Restore and enhance faces in an image with the Brainiall face-restoration engine. Detects all faces via RetinaFace, restores quality (fixes blur, noise, compression artifacts), and pastes them back. Optionally enhances the background with the Brainiall image-upscaling engine. GPU-accelerated, sub-3s latency. Args: image_base64: Base64-encoded image data containing faces (PNG, JPEG, WebP). upscale: Output upscale factor -- 1 to 4 (default: 2). enhance_background: Whether to enhance background with the Brainiall image-upscaling engine (default: true). Returns: dict with keys: - image (str): Base64-encoded restored image - format (str): Output image format - width (int): Output width - height (int): Output height - upscale (int): Scale factor applied - processing_time_ms (float): Processing time in milliseconds
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image_base64'], 'properties': {'upscale': {'type': 'integer', 'default': 2, 'description': 'Output upscale factor: 1-4 (default: 2)'}, 'image_base64': {'type': 'string', 'maxLength': 20000000, 'description': 'Base64-encoded image data containing one or more faces.'}, 'enhance_background': {'type': 'boolean', 'default': True, 'description': 'Enhance background with the Brainiall image-upscaling engine (default: true)'}}}
run_skillsets
Skillsets Enrichment Pipeline
Run a multi-skill enrichment pipeline over a document image or text in one call. Brainiall Skillsets engine. Returns per-skill outputs ready for indexing or RAG.
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'properties': {'text': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'default': None, 'description': 'Pre-extracted text (skip OCR)'}, 'image': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'default': None, 'description': 'Base64 image (triggers OCR)'}, 'skills': {'anyOf': [{'type': 'array', 'items': {'type': 'string'}}, {'type': 'null'}], 'default': None, 'description': 'Enrichment skills: ocr | entities | language | keyphrases | sentiment'}}}
understand_content
Multimodal Content Understanding
Multimodal extraction. Send an image, text, or both; define your schema of fields; get structured JSON. Brainiall Content Understanding engine. Unified multimodal field extraction over images and text.
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'properties': {'text': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'default': None, 'description': 'Optional pre-extracted text'}, 'image': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'default': None, 'description': 'Optional base64 image (will OCR first)'}, 'field_schema': {'anyOf': [{'type': 'object', 'additionalProperties': True}, {'type': 'null'}], 'default': None, 'description': 'Map of field_name -> description, e.g. {"invoice_id":"invoice number","total":"amount due"}'}}}
upscale_image
Upscale Image
Upscale image resolution with the Brainiall image-upscaling engine. Enhances image resolution by 2x or 4x with the GPU-accelerated Brainiall image-upscaling engine super-resolution. Processes in tiles (256x256) to manage VRAM. Maximum output dimension: 8192x8192. Args: image_base64: Base64-encoded image data (PNG, JPEG, or WebP). scale: Upscale factor -- 2 or 4 (default: 4). Returns: dict with keys: - image (str): Base64-encoded upscaled image - format (str): Output image format - width (int): Output width - height (int): Output height - scale (int): Scale factor applied - processing_time_ms (float): Processing time in milliseconds
Nur Lesen Externer Zugriff Idempotent
Eingabeschema
{'type': 'object', 'required': ['image_base64'], 'properties': {'scale': {'type': 'integer', 'default': 4, 'description': 'Upscale factor: 2 or 4 (default: 4)'}, 'image_base64': {'type': 'string', 'maxLength': 20000000, 'description': 'Base64-encoded image data. Supports PNG, JPEG, and WebP formats.'}}}
Hinzugefügt
understand_content
21. September 2026 02:47
Hinzugefügt
run_skillsets
21. September 2026 02:47
Hinzugefügt
document_to_markdown
21. September 2026 02:47
Hinzugefügt
document_tables
21. September 2026 02:47
Hinzugefügt
document_query
21. September 2026 02:47
Hinzugefügt
document_extract
21. September 2026 02:47
Hinzugefügt
check_image_service
21. September 2026 02:47
Hinzugefügt
restore_face
21. September 2026 02:47
Hinzugefügt
upscale_image
21. September 2026 02:47
Hinzugefügt
remove_background
21. September 2026 02:47