LLM Cost Estimator runs on your own machine — the client starts it, so there is no endpoint to ping. 29 installs a week from npm. Last commit 21 Jun 2026.
Token counting & multi-model LLM cost estimates: GPT-4o, Claude, Gemini, 25+. No API key.
Today is the operative word: we check LLM Cost Estimator every 15 minutes and re-read its code on every release. Watch it and you find out the day that stops being true.
This server runs on your own machine — install it with the package manager and the client starts it for you. Package name taken from the official registry entry.
claude mcp add llm-cost-estimator -- npx -y llm-cost-estimator-mcp
{
"mcpServers": {
"llm-cost-estimator": {
"args": [
"-y",
"llm-cost-estimator-mcp"
],
"command": "npx"
}
}
}
[mcp_servers.llm-cost-estimator]
command = "npx"
args = ["-y", "llm-cost-estimator-mcp"]
{
"mcpServers": {
"llm-cost-estimator": {
"args": [
"-y",
"llm-cost-estimator-mcp"
],
"command": "npx"
}
}
}
{
"mcpServers": {
"llm-cost-estimator": {
"args": [
"-y",
"llm-cost-estimator-mcp"
],
"command": "npx"
}
}
}
Access 100+ LLMs with one API: GPT-4, Claude, Gemini, Mistral, and more.
Multi-model AI debates: GPT-4o, Claude, Gemini & 200+ models discuss, then synthesize insight.
Analyze LLM API costs: token waste detection, caching savings estimates, model comparison
Live LLM API pricing: token prices, comparisons, cheapest-model lookups. No key required.
Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen
Cut LLM token costs: count tokens, estimate cost, slim prompts, and pick the cheapest capable model.
Other AI CLIs (Gemini, GPT, Claude, opencode, Ollama) as a council — ban-safe, no API keys.
Auto-route AI requests to 13 models (Claude, GPT, Gemini, Qwen, DeepSeek). 57% cost savings.
Answers built from our own checks of this server.