What this MCP does
Estimates whether local language models fit on specified GPUs or Apple Silicon systems based on memory and context requirements.
Tools
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['model'], 'properties': {'ctx': {'type': 'integer', 'minimum': 1024, 'description': 'Alias of context_tokens â\x80\x94 accepted because the REST API uses this name. Do not pass both with different values.'}, 'gpu': {'type': 'string', 'description': 'GPU name, fuzzy â\x80\x94 e.g. "RTX 4090", "RX 7900 XTX", "A100 80GB". Multi-GPU rigs: join with + â\x80\x94 e.g. "RTX 5090 + RTX 3090" (VRAM pools across cards). Provide gpu OR mac_ram_gb.'}, 'model': {'type': 'string', 'description': 'LLM name, fuzzy â\x80\x94 e.g. "GLM-4.7-Flash", "gpt-oss-20b", "gemma 31b"'}, 'quant': {'type': 'string', 'description': 'Weight quantization. GPU: Q4_K_M(default)/Q5_K_M/Q6_K/Q8_0/FP16. Mac: 4/8(default)/16 (bits).'}, 'kv_bits': {'enum': [16, 8, 4], 'type': 'number', 'description': 'KV-cache quantization bits (default 16 = F16)'}, 'gpu_count': {'type': 'integer', 'maximum': 8, 'minimum': 1, 'description': 'Number of identical copies of the gpu (e.g. gpu="RTX 3090", gpu_count=2 for a 2Ã\x973090 rig). Default 1.'}, 'mac_ram_gb': {'type': 'integer', 'maximum': 2048, 'minimum': 8, 'description': 'Apple Silicon unified memory in GB â\x80\x94 e.g. 16, 64, 512. Provide gpu OR mac_ram_gb.'}, 'context_tokens': {'type': 'integer', 'minimum': 1024, 'description': 'Context length in tokens (default 8192). Alias: ctx (same field as the REST API).'}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'properties': {}}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'properties': {'gpu': {'type': 'string', 'description': 'GPU name, fuzzy. Multi-GPU rigs: join with + (e.g. "RTX 5090 + RTX 3090"). Provide gpu OR mac_ram_gb.'}, 'gpu_count': {'type': 'integer', 'maximum': 8, 'minimum': 1, 'description': 'Number of identical copies of the gpu. Default 1.'}, 'mac_ram_gb': {'type': 'integer', 'maximum': 2048, 'minimum': 8, 'description': 'Apple Silicon unified memory GB. Provide gpu OR mac_ram_gb.'}}, 'additionalProperties': False}
Recent tool changes
Similar MCP servers
TinyFn
Offers deterministic utility functions for mathematics, conversions, validation, hashing, encoding, arrays, dates, colors, and pa…
hyperion
Acts as a paid MCP tool marketplace and utility gateway with server discovery, HTTP and JavaScript tools, research, data conversi…
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
Vee3
Manages Clerk authentication infrastructure, including users, organizations, domains, sessions, tokens, OAuth, SSO, machines, per…
Courier
Provides notification delivery infrastructure for users, tenants, lists, templates, preferences, journeys, automations, brands, a…
GripForge
Generates, rigs, animates, analyzes, and packages game characters, weapons, armor, effects, environments, and engine-ready assets.
IA-QA — 130+ QA & Dev Tools for AI Agents
Provides deterministic QA, evaluation, testing, code analysis, prompt and RAG checks, model comparison, and web security diagnost…
apis-io
Provides catalog search, comparison, scoring, enrichment, lists, and dataset exports for APIs, providers, specifications, workflo…