Serveur MCP

InferIndex

io.github.InferIndex/inferindex
Cloud et infrastructure Outils développeur Public et accessible MCP 2026-07-28

Ce que fait ce MCP

Compares LLM API providers, prices, reliability, historical pricing, and estimated workload costs.

cheapest
Cheapest offers for a model
Cheapest current API offers for one model across direct providers and aggregators, in USD per 1M tokens (input, output, blended 3:1). Stale prices, and flex/batch tiers, are excluded by default. Optional filters (context, tools, JSON, vision, region, no training on prompts, open sign-up) and usage (tokens per request, requests per day) to get an estimated cost per request and per month. Returns the winner and the first offers.
Lecture seule Idempotent
Schéma d’entrée
{'type': 'object', 'required': ['model'], 'properties': {'json': {'type': 'boolean', 'description': 'Only offers that support JSON output'}, 'limit': {'type': 'integer', 'maximum': 25, 'minimum': 1, 'description': 'Number of offers to return (default 5, max 25)'}, 'model': {'type': 'string', 'description': "Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure."}, 'tools': {'type': 'boolean', 'description': 'Only offers that support tool calling'}, 'region': {'type': 'string', 'description': 'Only providers that process data in this region: eu, us, …'}, 'strict': {'type': 'boolean', 'description': 'Exclude offers whose provider does not publish the filtered information (by default they are kept and flagged)'}, 'vision': {'type': 'boolean', 'description': 'Only offers that accept image input'}, 'min_context': {'type': 'integer', 'minimum': 0, 'description': 'Minimum context window in tokens'}, 'no_training': {'type': 'boolean', 'description': 'Only providers whose published terms say they do not train on your prompts'}, 'no_waitlist': {'type': 'boolean', 'description': 'Only providers with open sign-up (no waitlist or invitation)'}, 'cached_ratio': {'type': 'number', 'maximum': 1, 'minimum': 0, 'description': "Share of input tokens served from the provider's prompt cache (0 to 1)"}, 'include_tiers': {'type': 'string', 'description': 'Also include lower-priority service tiers, comma-separated: flex, batch (hidden by default)'}, 'output_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Output tokens per request, for the estimated cost'}, 'prompt_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Input tokens per request, for the estimated cost'}, 'requests_per_day': {'type': 'integer', 'minimum': 0, 'description': 'Requests per day, to also get an estimated monthly cost'}}, 'additionalProperties': False}
compare_providers
Compare providers for a model
Current offers for one model, one line per provider and source (direct or via an aggregator), cheapest first (10 by default), with price, context, quantization, published conditions (training on prompts, data regions, sign-up) and reliability from official status pages.
Lecture seule Idempotent
Schéma d’entrée
{'type': 'object', 'required': ['model'], 'properties': {'sort': {'enum': ['blended', 'input', 'output', 'estimated_cost'], 'type': 'string', 'description': 'Sort order (default blended); estimated_cost needs prompt_tokens or output_tokens'}, 'limit': {'type': 'integer', 'maximum': 50, 'minimum': 1, 'description': 'Number of offers to return (default 10, max 50)'}, 'model': {'type': 'string', 'description': "Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure."}, 'region': {'type': 'string', 'description': 'Only providers that process data in this region: eu, us, …'}, 'strict': {'type': 'boolean', 'description': 'Exclude offers whose provider does not publish the filtered information (by default they are kept and flagged)'}, 'no_training': {'type': 'boolean', 'description': 'Only providers whose published terms say they do not train on your prompts'}, 'no_waitlist': {'type': 'boolean', 'description': 'Only providers with open sign-up (no waitlist or invitation)'}, 'cached_ratio': {'type': 'number', 'maximum': 1, 'minimum': 0, 'description': "Share of input tokens served from the provider's prompt cache (0 to 1)"}, 'include_tiers': {'type': 'string', 'description': 'Also include lower-priority service tiers, comma-separated: flex, batch (hidden by default)'}, 'output_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Output tokens per request, for the estimated cost'}, 'prompt_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Input tokens per request, for the estimated cost'}, 'requests_per_day': {'type': 'integer', 'minimum': 0, 'description': 'Requests per day, to also get an estimated monthly cost'}}, 'additionalProperties': False}
estimate_cost
Estimate the cost of a workload
Estimated cost of a workload on one model at each provider: cost per request, and per month if requests_per_day is given, taking the provider's tiered pricing and prompt-cache price into account. Offers sorted by estimated cost, cheapest first.
Lecture seule Idempotent
Schéma d’entrée
{'type': 'object', 'required': ['model'], 'properties': {'limit': {'type': 'integer', 'maximum': 25, 'minimum': 1, 'description': 'Number of offers to return (default 5, max 25)'}, 'model': {'type': 'string', 'description': "Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure."}, 'cached_ratio': {'type': 'number', 'maximum': 1, 'minimum': 0, 'description': "Share of input tokens served from the provider's prompt cache (0 to 1)"}, 'output_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Output tokens per request, for the estimated cost'}, 'prompt_tokens': {'type': 'integer', 'minimum': 0, 'description': 'Input tokens per request, for the estimated cost'}, 'requests_per_day': {'type': 'integer', 'minimum': 0, 'description': 'Requests per day, to also get an estimated monthly cost'}}, 'additionalProperties': False}
price_history
Price history of a model
Price history of one model: every offer tracked by InferIndex (daily or weekly min/max/last price in USD, or raw price changes), plus the official price of the model's lab over time. Give either days, or from/to (YYYY-MM-DD), or at (a date) for the prices in effect that day.
Lecture seule Idempotent
Schéma d’entrée
{'type': 'object', 'required': ['model'], 'properties': {'at': {'type': 'string', 'description': 'A single date, YYYY-MM-DD: prices in effect that day'}, 'to': {'type': 'string', 'description': 'End date, YYYY-MM-DD'}, 'days': {'type': 'integer', 'minimum': 1, 'description': 'Number of days back from today (default 7)'}, 'from': {'type': 'string', 'description': 'Start date, YYYY-MM-DD (with to, instead of days)'}, 'limit': {'type': 'integer', 'maximum': 500, 'minimum': 1, 'description': 'Maximum number of points (default 100, max 500)'}, 'model': {'type': 'string', 'description': "Model id or name, e.g. 'deepseek-v3.2', 'deepseek/deepseek-v4-pro', 'gpt-5.6-luna'. Use search_models when unsure."}, 'provider': {'type': 'string', 'description': 'Only this provider'}, 'granularity': {'enum': ['day', 'week', 'raw'], 'type': 'string', 'description': 'day (default), week, or raw price changes'}}, 'additionalProperties': False}
search_models
Search models
Find the exact id of an LLM tracked by InferIndex from a name or partial name (e.g. 'deepseek', 'qwen3 max', 'claude opus'). Returns matching model ids and names, best match first.
Lecture seule Idempotent
Schéma d’entrée
{'type': 'object', 'required': ['query'], 'properties': {'query': {'type': 'string', 'description': 'Model name or part of it'}}, 'additionalProperties': False}
Ajouté
estimate_cost
19 September 2026 02:40
Ajouté
price_history
19 September 2026 02:40
Ajouté
compare_providers
19 September 2026 02:40
Ajouté
cheapest
19 September 2026 02:40
Ajouté
search_models
19 September 2026 02:40