agent-trace-auditor
Qué hace este MCP
Audits and compares AI agent execution traces for loops, tool errors, schema violations, cost overruns, regressions, and latency issues.
Herramientas
Esquema de entrada
{'type': 'object', 'required': ['trace'], 'properties': {'trace': {'type': 'array', 'items': {'type': 'object', 'properties': {'ts': {'type': 'string'}, 'tool': {'type': 'string'}, 'error': {}, 'input': {}, 'model': {'type': 'string'}, 'output': {}, 'latency_ms': {'type': 'number'}}}, 'description': 'Agent execution steps. Each step: {tool, input, output, model, ts, latency_ms, error}'}, 'budget_usd': {'type': 'number', 'description': 'Cost cap in USD (default 5.0)'}, 'tool_schemas': {'type': 'object', 'description': 'Per-tool required field schemas for input validation'}}}
Esquema de entrada
{'type': 'object', 'required': ['trace_a', 'trace_b'], 'properties': {'trace_a': {'type': 'array', 'description': 'Baseline (previous) trace'}, 'trace_b': {'type': 'array', 'description': 'New trace to compare against baseline'}, 'budget_usd': {'type': 'number'}}}
Esquema de entrada
{'type': 'object', 'required': ['trace'], 'properties': {'trace': {'type': 'array', 'description': 'Agent execution steps'}, 'budget_usd': {'type': 'number'}}}
Cambios recientes en herramientas
Servidores MCP similares
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
IA-QA — 130+ QA & Dev Tools for AI Agents
Provides deterministic QA, evaluation, testing, code analysis, prompt and RAG checks, model comparison, and web security diagnost…
CompletionKit
Runs prompt evaluation workflows over datasets using deterministic checks and LLM judges, with metrics, scoring runs, agreements,…
Replicate
Provides access to Replicate models, versions, collections, hardware, predictions, and deployments for running and managing hoste…
Huggingface
Provides access to Hugging Face model, dataset, and Space metadata, alongside broader structured research and data-routing tools.
Ai Model Experiments
Runs prompts across multiple AI models and compares their outputs, costs, latency, token usage, and errors.
rubrkit
Manages AI evaluation artifacts, rubric audits, eval runs, golden cases, proof reports, and drift monitoring for LLM outputs.
Hipocampo MCP
Provides bilingual semantic memory, embeddings, graph links, code indexing and search, context preload, deduplication, compressio…