Inference AIops runs on your own machine — the client starts it, so there is no endpoint to ping. 887 installs a week from pypi. Last commit 3 Aug 2026.
Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.
This server runs on your own machine — install it with the package manager and the client starts it for you. Package name taken from the official registry entry.
claude mcp add inference-aiops -- uvx inference-aiops
{
"mcpServers": {
"inference-aiops": {
"args": [
"inference-aiops"
],
"command": "uvx"
}
}
}
[mcp_servers.inference-aiops]
command = "uvx"
args = ["inference-aiops"]
{
"mcpServers": {
"inference-aiops": {
"args": [
"inference-aiops"
],
"command": "uvx"
}
}
}
{
"mcpServers": {
"inference-aiops": {
"args": [
"inference-aiops"
],
"command": "uvx"
}
}
}
Answers built from our own checks of this server.