mcpbeat Sign in

LMX Cloud LLM Inference MCP Server

answering

LMX Cloud LLM Inference is answering right now. Last checked 5 min ago. It exposes 9 tools.

OpenAI-compatible LLM MCP (7 tools); chat via balance key or x402 USDC on Base

Uptime history 47 days of history · worst day 99%
47 days agonow
100.0%
Uptime 24h
92 of 92 checks
9
Tools
read from the server
390 ms
Response time
average over 24h
open, no key
Access
streamable-http

What changed 1

Every tool that appeared, vanished or quietly changed what it asks for. Recorded since 11 August 2026. No other catalogue keeps this.

11 Aug a tool appeared extract_pdf

Nothing serious here today

Today is the operative word: we check LMX Cloud LLM Inference every 15 minutes and re-read its code on every release. Watch it and you find out the day that stops being true.

Three servers free · no card

Connect this server

Endpoint below is the one we actually reach during checks — not the one copied from a README. Last verified 5 min ago.

run in your terminal
claude mcp add mcp-server --transport http https://mcp.lmxcloud.io/mcp
~/Library/Application Support/Claude/claude_desktop_config.json
{
  "mcpServers": {
    "mcp-server": {
      "url": "https://mcp.lmxcloud.io/mcp"
    }
  }
}
~/.codex/config.toml
[mcp_servers.mcp-server]
url = "https://mcp.lmxcloud.io/mcp"
.cursor/mcp.json
{
  "mcpServers": {
    "mcp-server": {
      "url": "https://mcp.lmxcloud.io/mcp"
    }
  }
}
.vscode/mcp.json
{
  "mcpServers": {
    "mcp-server": {
      "url": "https://mcp.lmxcloud.io/mcp"
    }
  }
}

Available tools 9

Read directly from the server with tools/list, grouped by what they act on. If a tool disappears, we record the date.

balance
get_balance
Fetch USD credit balance for the caller API key. Requires authentication.
chat
chat_completion
Call LMX Cloud OpenAI-compatible chat completions. Prefers a pre-funded API key (Bearer / api_key); when omitted and x402 is enabled, requires a USDC pay-per-call payment. Optional image_url / images enable vision input (OpenAI content-parts format) — use a vision model such as llama-3.2-90b-vision, qwen-3.6-35b, or qwen-3.5-35b.
extract
extract_pdf
Extract text and light structure (title, headings, page count) from a PDF via LMX pdf-extract. Provide file_url and/or file_base64. Requires a real LMX API key (same gate as web_search).
models
list_models
List currently supported LMX model aliases and providers.
pricing
get_pricing
Fetch current LMX Cloud per-call pricing catalog.
quote
quote_price
Estimate USDC cost for a single model call. Uses GET /v1/pricing with model and token params.
status
get_status
Fetch LMX Cloud provider health, fallback chain, and anchoring status.
usage
get_usage
Fetch request and token usage totals for the caller API key. Requires authentication.
web
web_search
Real-time web search via LMX (Brave Search passthrough). Fixed per-call USDC price from the caller's API key balance. Returns title/url/snippet results.

Endpoints

URLTransportStateLatencyChecked
https://mcp.lmxcloud.io/mcp streamable-http answering 290 ms 5 min ago

Alternatives to LMX Cloud LLM Inference

same job, measured the same way
Myai MCP
by myaitoken

Decentralized AI inference on Base: OpenAI-compatible chat via community GPUs, paid in MYAI

87 installs/wk local only
BridgeNode MCP
by applefanaimail-blip

BridgeNode — x402 pay-per-request AI inference. OpenAI-compatible API + MCP server, Solana USDC.

3 tools answering
BridgeNode MCP
by bridgenode-ai

BridgeNode — x402 pay-per-request AI inference. OpenAI-compatible API + MCP, Solana USDC, gas-free.

416 installs/wk 3 tools answering
C
Pch X402 MCP
by pathcoursehealth

PathCourse Health inference SKUs as x402 paid MCP tools, billed per call in USDC on Base.

42 installs/wk local only
S
Otto Intel
by ottoai

Otto AI hosted MCP: 16 pay-per-call market-intelligence tools, USDC via x402 on Base, no API keys.

32 tools answering
Bay Run
by barneywohl

Free OpenAI-compatible inference with signed provenance receipts and 3 focused MCP tools.

3 tools answering
Hubris
by hubris

OpenAI-compatible LLM gateway for Russia: 500+ models, ruble pricing, balance, chat.

47 installs/wk answering
Scrape402
by scrape402

Pay-per-call agent tools via x402 (USDC on Base): chat, prices, funding, RNG. No account or keys.

5 tools answering

LMX Cloud LLM Inference — questions

Answers built from our own checks of this server.

What can LMX Cloud LLM Inference do?
It exposes 9 tools, read directly from the server on our last check. Among them: chat_completion, extract_pdf, get_balance, get_pricing, get_status, get_usage and 3 more. The full list with descriptions is on this page — we take it from the server itself via tools/list, not from a README. How MCP servers expose tools in the first place →
Is LMX Cloud LLM Inference working right now?
We send a real MCP handshake every 15 minutes. Over the last 24 hours 92 of 92 checks got a reply (100.0%), average response time 390 ms. The bar chart above shows every period we have measured.
How do I connect LMX Cloud LLM Inference?
Copy the ready config from this page — we generate it for Claude Code, Claude Desktop, Codex, Cursor and VS Code, each with the file path that client actually reads. It is a remote server, so there is nothing to install — the client connects to the address.
Does LMX Cloud LLM Inference need an API key?
No. LMX Cloud LLM Inference completed a full MCP handshake with us as an anonymous client and listed its tools without asking for anything. All 9 of them are readable on this page. This is what we observed, not what the docs claim.
How fast is LMX Cloud LLM Inference?
It answers our handshake in 390 ms on average, which is faster than 39% of all working MCP servers we measure. The comparison comes from our own checks across the whole registry, every 15 minutes.