llm-output-quality-monitor
Qué hace este MCP
Checks LLM responses for consistency, quality, schema compliance, drift, and heuristic hallucination risk.
Herramientas
Esquema de entrada
{'type': 'object', 'required': ['responses'], 'properties': {'responses': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Array of responses to compare'}}}
Esquema de entrada
{'type': 'object', 'required': ['currentResponse', 'previousResponse'], 'properties': {'threshold': {'type': 'number', 'description': 'Drift threshold (0-1, default: 0.15)'}, 'currentResponse': {'type': 'string', 'description': 'Current LLM response'}, 'previousResponse': {'type': 'string', 'description': 'Previous LLM response'}}}
Esquema de entrada
{'type': 'object', 'required': ['response'], 'properties': {'context': {'type': 'string', 'description': 'Reference context for grounding'}, 'response': {'type': 'string', 'description': 'LLM response to analyze'}}}
Esquema de entrada
{'type': 'object', 'required': ['response'], 'properties': {'response': {'type': 'string', 'description': 'LLM response to validate'}, 'maxLength': {'type': 'number', 'description': 'Maximum response length (default: 10000)'}, 'minLength': {'type': 'number', 'description': 'Minimum response length (default: 10)'}, 'strictFormat': {'type': 'boolean', 'description': 'Enforce punctuation and capitalization'}}}
Esquema de entrada
{'type': 'object', 'required': ['response', 'schema'], 'properties': {'schema': {'type': 'object', 'description': 'JSON schema definition'}, 'response': {'type': 'string', 'description': 'JSON response to validate'}}}
Cambios recientes en herramientas
Servidores MCP similares
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
IA-QA — 130+ QA & Dev Tools for AI Agents
Provides deterministic QA, evaluation, testing, code analysis, prompt and RAG checks, model comparison, and web security diagnost…
CompletionKit
Runs prompt evaluation workflows over datasets using deterministic checks and LLM judges, with metrics, scoring runs, agreements,…
Replicate
Provides access to Replicate models, versions, collections, hardware, predictions, and deployments for running and managing hoste…
Huggingface
Provides access to Hugging Face model, dataset, and Space metadata, alongside broader structured research and data-routing tools.
Ai Model Experiments
Runs prompts across multiple AI models and compares their outputs, costs, latency, token usage, and errors.
rubrkit
Manages AI evaluation artifacts, rubric audits, eval runs, golden cases, proof reports, and drift monitoring for LLM outputs.
Hipocampo MCP
Provides bilingual semantic memory, embeddings, graph links, code indexing and search, context preload, deduplication, compressio…