MCPサーバー

ocr

com.auto-reader/ocr

このMCPでできること

Performs multilingual OCR, document field extraction, image and manga reading, translation, and Saudi ZATCA invoice QR processing.

create_api_key
Provision a new Auto-Reader OCR API key instantly, with no human steps. Pass an optional email to unlock the larger free tier (about 250 credits/day, vs about 25/day for an email-less trial key). Store the returned key and pass it as api_key on future calls.
入力スキーマ
{'type': 'object', 'required': [], 'properties': {'email': {'type': 'string', 'description': 'Optional email to attach for the larger free tier and a verification link.'}}}
extract_document
Extract STRUCTURED FIELDS from a document image: invoices, receipts, ID cards — or any custom JSON schema you supply. Every field returns {value, confidence, box} where the confidence and box come from the OCR geometry (never model guesswork); absent fields are null. preset="zatca" additionally decodes the Saudi ZATCA e-invoice QR (TLV) and cross-validates it against the printed fields — use it for Saudi tax invoices. Arabic-first accuracy. 5 credits/page (zatca 7).
入力スキーマ
{'type': 'object', 'required': ['image_base64'], 'properties': {'lang': {'type': 'string', 'default': 'auto', 'description': 'Language hint; default auto.'}, 'preset': {'enum': ['invoice', 'receipt', 'id', 'zatca'], 'type': 'string', 'description': 'Built-in schema. Use zatca for Saudi e-invoices (adds QR validation).'}, 'schema': {'type': 'object', 'description': 'Custom extraction schema instead of a preset: an object whose keys are the fields you want, values describing them, e.g. {"policy_number": "string|null"}.'}, 'api_key': {'type': 'string', 'description': 'Optional Auto-Reader OCR key (nsk_live_...). If omitted, a free trial key is auto-provisioned and returned to you in the result.'}, 'image_base64': {'type': 'string', 'description': 'The document image as base64 (data: URI prefix accepted).'}}}
get_usage
Check your Auto-Reader OCR key: tier, remaining daily free credits, prepaid credit balance, subscription allowance, and per-minute rate limit. Use it to throttle yourself before hitting a limit.
入力スキーマ
{'type': 'object', 'required': [], 'properties': {'api_key': {'type': 'string', 'description': 'Optional key (nsk_live_...). Auto-provisioned if omitted.'}}}
ocr_and_translate
One call: OCR an image, then translate every line into target_lang. Arabic-first OCR and manga-aware Japanese with right-to-left-aware layout, followed by LLM translation. Automatic source-language detection. Provide the image as base64. Ideal for reading foreign documents, signs, manga, or receipts end-to-end in a single step.
入力スキーマ
{'type': 'object', 'required': ['image_base64', 'target_lang'], 'properties': {'api_key': {'type': 'string', 'description': 'Optional key (nsk_live_...). Auto-provisioned if omitted.'}, 'target_lang': {'type': 'string', 'description': 'Language to translate into, as a name or code (e.g. English, ar, ja).'}, 'image_base64': {'type': 'string', 'description': 'The image encoded as base64 (a data: URI prefix is accepted and stripped).'}}}
ocr_image
Extract text from an image with GPU OCR. Best-in-class Arabic (plus Persian/Urdu) accuracy, manga-aware vertical Japanese, and strong English, French, Spanish, German, Chinese, Korean, Russian, Italian and Portuguese — 13+ languages. Automatic language and script detection with lang="auto". Returns reading-order layout text (right-to-left aware, paragraph-gapped) that is ready to feed an LLM or show a human, plus the detected language, the engine used, and the number of text blocks found. Provide the image as base64. Use the mode hint (document | receipt | manga | scene) to tune detection.
入力スキーマ
{'type': 'object', 'required': ['image_base64'], 'properties': {'lang': {'type': 'string', 'default': 'auto', 'description': 'Language/script hint. Default "auto" detects it. Codes: ar, fa, ur, en, fr, es, de, ja, zh, ko, ru, it, pt.'}, 'mode': {'enum': ['document', 'receipt', 'manga', 'scene'], 'type': 'string', 'default': 'document', 'description': 'Content hint that tunes detection and prompts. Default "document".'}, 'api_key': {'type': 'string', 'description': 'Optional Auto-Reader OCR key (nsk_live_...). If omitted, a free trial key is auto-provisioned and returned to you in the result.'}, 'quality': {'enum': ['standard', 'high'], 'type': 'string', 'default': 'standard', 'description': '"standard" (default) lets a confidence gate decide whether the vision model re-reads the page. "high" always re-reads it — use when accuracy matters more than cost or latency (costs 2 extra credits and adds a few seconds). You are charged the extra ONLY when it actually applies: check quality_applied in the result, and notice tells you why if it is false (receipt mode, a manga-engine page, an out-of-scope language, or the vision read failing its quality guards).'}, 'image_base64': {'type': 'string', 'description': 'The image encoded as base64 (a data: URI prefix is accepted and stripped).'}}}
read_manga
OCR a comic/manga page, routed by LANGUAGE (not the blanket "manga = Japanese" assumption). Japanese goes to the manga specialist reader that reads vertical, hand-lettered speech bubbles in right-to-left order; Korean manhwa, Chinese manhua and other scripts use their own OCR pack; low-confidence pages escalate to the vision model. Returns text blocks in reading order plus the detected language, the engine used, and whether the page is vertical. Provide the image as base64. Pass lang explicitly (ko/zh/...) for the best non-Japanese result; default "auto" detects it.
入力スキーマ
{'type': 'object', 'required': ['image_base64'], 'properties': {'lang': {'type': 'string', 'default': 'auto', 'description': 'Language hint. Default "auto". Codes: ja (manga specialist), ko, zh, ar, en, fr, es, de, ru, it, pt.'}, 'api_key': {'type': 'string', 'description': 'Optional Auto-Reader OCR key (nsk_live_...). If omitted, a free trial key is auto-provisioned and returned to you in the result.'}, 'image_base64': {'type': 'string', 'description': 'The manga/comic page as base64 (a data: URI prefix is accepted and stripped).'}}}
translate_text
Translate text between 13+ languages with an LLM. Arabic-first quality, with formality control (formal/informal) and optional context to disambiguate meaning. Handles both short dictionary-style word lookups and full documents. Returns the translation and, when available, alternative phrasings.
入力スキーマ
{'type': 'object', 'required': ['text', 'target_lang'], 'properties': {'text': {'type': 'string', 'description': 'The text to translate.'}, 'api_key': {'type': 'string', 'description': 'Optional key (nsk_live_...). Auto-provisioned if omitted.'}, 'context': {'type': 'string', 'description': 'Optional background text that improves accuracy (it is not translated).'}, 'formality': {'enum': ['formal', 'informal'], 'type': 'string', 'description': 'Optional register for the output.'}, 'target_lang': {'type': 'string', 'description': 'Target language, as a name or code (e.g. English, Arabic, ar, ja, fr).'}}}
追加
create_api_key
2026年9月17日12:33
追加
get_usage
2026年9月17日12:33
追加
ocr_and_translate
2026年9月17日12:33
追加
translate_text
2026年9月17日12:33
追加
read_manga
2026年9月17日12:33
追加
ocr_image
2026年9月17日12:33
追加
extract_document
2026年9月17日12:33