このMCPでできること
Provides hosted speech-to-text transcription and speech emotion or tone analysis.
ツール
入力スキーマ
{'type': 'object', 'properties': {'model': {'enum': ['oruk-resonance', 'oruk-fourier'], 'type': 'string', 'description': 'oruk-resonance (full local pipeline: transcription, emotion, style, affect, analysis) or oruk-fourier (parallel transcript and native 15-label emotion, with the shared 16-label style model). Resonance is the default.'}, 'detail': {'enum': ['compact', 'full'], 'type': 'string', 'description': 'compact (default) returns top label scores and condensed segments; full preserves all returned labels, segments, and word-level timings, subject to response-size limits. It does not expose unreturned label scores.'}, 'api_key': {'type': 'string', 'maxLength': 200, 'description': 'Only for temporary keys from oruk_create_trial_key. Permanent keys belong in your MCP client config as an "Authorization: Bearer <key>" header, never in tool arguments.'}, 'diarize': {'type': 'boolean', 'description': 'Label speakers (oruk-resonance only; the model is switched to oruk-resonance automatically). Speaker diarization locates turns, then Resonance scores each turn with its own text, emotions, and styles. Use for calls, meetings, and interviews. Included in subscription plan minutes. Processing details: https://oruk.ai/security#processing.'}, 'filename': {'type': 'string', 'maxLength': 160, 'description': 'Original filename including extension (e.g. call.wav). Helps decoding when audio_base64 is used.'}, 'audio_url': {'type': 'string', 'format': 'uri', 'maxLength': 2000, 'description': 'Publicly fetchable audio file URL (wav, flac, mp3, m4a, ogg, webm; up to 30 MB / 60 minutes of English speech).'}, 'audio_base64': {'type': 'string', 'maxLength': 11500000, 'description': 'Base64-encoded audio bytes for local files (up to 8 MiB decoded). Prefer audio_url for anything larger.'}}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {'model': {'enum': ['oruk-resonance', 'oruk-fourier'], 'type': 'string', 'description': 'oruk-resonance (full local pipeline: transcription, emotion, style, affect, analysis) or oruk-fourier (parallel transcript and native 15-label emotion, with the shared 16-label style model). Resonance is the default.'}, 'detail': {'enum': ['compact', 'full'], 'type': 'string', 'description': 'compact (default) returns top label scores and condensed segments; full preserves all returned labels, segments, and word-level timings, subject to response-size limits. It does not expose unreturned label scores.'}, 'api_key': {'type': 'string', 'maxLength': 200, 'description': 'Only for temporary keys from oruk_create_trial_key. Permanent keys belong in your MCP client config as an "Authorization: Bearer <key>" header, never in tool arguments.'}, 'diarize': {'type': 'boolean', 'description': 'Label speakers (oruk-resonance only; the model is switched to oruk-resonance automatically). Speaker diarization locates turns, then Resonance scores each turn with its own text, emotions, and styles. Use for calls, meetings, and interviews. Included in subscription plan minutes. Processing details: https://oruk.ai/security#processing.'}, 'filename': {'type': 'string', 'maxLength': 160, 'description': 'Original filename including extension (e.g. call.wav). Helps decoding when audio_base64 is used.'}, 'audio_url': {'type': 'string', 'format': 'uri', 'maxLength': 2000, 'description': 'Publicly fetchable audio file URL (wav, flac, mp3, m4a, ogg, webm; up to 30 MB / 60 minutes of English speech).'}, 'audio_base64': {'type': 'string', 'maxLength': 11500000, 'description': 'Base64-encoded audio bytes for local files (up to 8 MiB decoded). Prefer audio_url for anything larger.'}}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {'api_key': {'type': 'string', 'maxLength': 200, 'description': 'Only for temporary keys from oruk_create_trial_key. Permanent keys belong in the Authorization header of your MCP client config.'}}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {}, 'additionalProperties': False}
入力スキーマ
{'type': 'object', 'properties': {'model': {'enum': ['oruk-resonance', 'oruk-fourier'], 'type': 'string', 'description': 'oruk-resonance (full local pipeline: transcription, emotion, style, affect, analysis) or oruk-fourier (parallel transcript and native 15-label emotion, with the shared 16-label style model). Resonance is the default.'}, 'detail': {'enum': ['compact', 'full'], 'type': 'string', 'description': 'compact (default) returns top label scores and condensed segments; full preserves all returned labels, segments, and word-level timings, subject to response-size limits. It does not expose unreturned label scores.'}, 'api_key': {'type': 'string', 'maxLength': 200, 'description': 'Only for temporary keys from oruk_create_trial_key. Permanent keys belong in your MCP client config as an "Authorization: Bearer <key>" header, never in tool arguments.'}, 'diarize': {'type': 'boolean', 'description': 'Label speakers (oruk-resonance only; the model is switched to oruk-resonance automatically). Speaker diarization locates turns, then Resonance scores each turn with its own text, emotions, and styles. Use for calls, meetings, and interviews. Included in subscription plan minutes. Processing details: https://oruk.ai/security#processing.'}, 'filename': {'type': 'string', 'maxLength': 160, 'description': 'Original filename including extension (e.g. call.wav). Helps decoding when audio_base64 is used.'}, 'audio_url': {'type': 'string', 'format': 'uri', 'maxLength': 2000, 'description': 'Publicly fetchable audio file URL (wav, flac, mp3, m4a, ogg, webm; up to 30 MB / 60 minutes of English speech).'}, 'audio_base64': {'type': 'string', 'maxLength': 11500000, 'description': 'Base64-encoded audio bytes for local files (up to 8 MiB decoded). Prefer audio_url for anything larger.'}}, 'additionalProperties': False}
最近のツール変更
類似のMCPサーバー
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
Social Fetch
Retrieves public content and metadata from social networks, music services, Facebook Marketplace, events, profiles, posts, commen…
Hermoso
Supports ad research, competitor analysis, video and static ad creation, campaign policy checks, content publishing, and performa…
ElevenLabs
Manages ElevenLabs speech, voices, pronunciation dictionaries, podcasts, dubbing, audio projects, workspaces, and related media r…
John's Essentials
Analyzes and converts documents, images, audio, video, archives, data files, and other media formats, with PDF chat and file insp…
MiOffice — AI-Powered Workspace Studio
Provides browser-based tools for processing PDFs, images, video, and audio, including generation, enhancement, conversion, transc…
BlitzReels Video Editor
Creates and edits short-form videos, including timelines, captions, transitions, AI-generated visuals, music, voiceovers, sound e…
Heista
Analyzes video advertising, builds brand intelligence profiles, researches markets, and generates advertising scripts and creativ…