Inference AIops
About
Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.
Use this server
Add it to your MCP client. The catalogue lists this config verbatim from its source — it does not run, download, or vouch for the server. Review the command before you run it.
Transportstdio (runs a local command)
Claude Code
claude mcp add inference-aiops -- uvx inference-aiops
Cursor
If the button doesn't open Cursor, use the JSON below — Cursor accepts the same mcpServers config.
VS Code
code --add-mcp '{"name":"inference-aiops","command":"uvx","args":["inference-aiops"]}'
JSON
{
"mcpServers": {
"inference-aiops": {
"command": "uvx",
"args": [
"inference-aiops"
]
}
}
}