此 MCP 可以做什么
Checks LLM responses for consistency, quality, schema compliance, drift, and heuristic hallucination risk.
工具
输入模式
{'type': 'object', 'required': ['responses'], 'properties': {'responses': {'type': 'array', 'items': {'type': 'string'}, 'description': 'Array of responses to compare'}}}
输入模式
{'type': 'object', 'required': ['currentResponse', 'previousResponse'], 'properties': {'threshold': {'type': 'number', 'description': 'Drift threshold (0-1, default: 0.15)'}, 'currentResponse': {'type': 'string', 'description': 'Current LLM response'}, 'previousResponse': {'type': 'string', 'description': 'Previous LLM response'}}}
输入模式
{'type': 'object', 'required': ['response'], 'properties': {'context': {'type': 'string', 'description': 'Reference context for grounding'}, 'response': {'type': 'string', 'description': 'LLM response to analyze'}}}
输入模式
{'type': 'object', 'required': ['response'], 'properties': {'response': {'type': 'string', 'description': 'LLM response to validate'}, 'maxLength': {'type': 'number', 'description': 'Maximum response length (default: 10000)'}, 'minLength': {'type': 'number', 'description': 'Minimum response length (default: 10)'}, 'strictFormat': {'type': 'boolean', 'description': 'Enforce punctuation and capitalization'}}}
输入模式
{'type': 'object', 'required': ['response', 'schema'], 'properties': {'schema': {'type': 'object', 'description': 'JSON schema definition'}, 'response': {'type': 'string', 'description': 'JSON response to validate'}}}
近期工具变更
类似的 MCP 服务器
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
IA-QA — 130+ QA & Dev Tools for AI Agents
Provides deterministic QA, evaluation, testing, code analysis, prompt and RAG checks, model comparison, and web security diagnost…
CompletionKit
Runs prompt evaluation workflows over datasets using deterministic checks and LLM judges, with metrics, scoring runs, agreements,…
Replicate
Provides access to Replicate models, versions, collections, hardware, predictions, and deployments for running and managing hoste…
Huggingface
Provides access to Hugging Face model, dataset, and Space metadata, alongside broader structured research and data-routing tools.
Ai Model Experiments
Runs prompts across multiple AI models and compares their outputs, costs, latency, token usage, and errors.
rubrkit
Manages AI evaluation artifacts, rubric audits, eval runs, golden cases, proof reports, and drift monitoring for LLM outputs.
Hipocampo MCP
Provides bilingual semantic memory, embeddings, graph links, code indexing and search, context preload, deduplication, compressio…