Back to Browse

Inference AIops MCP Server

AI & MLModerate5.2MCP RegistryLocal
Free

Server data from the Official MCP Registry

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

About

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

Security Report

5.2
Moderate5.2Moderate Risk

A well-architected governance harness for AI-ops inference clusters with strong audit, budget-guard, and undo-tracking mechanisms. Code is clean, authentication is optional but properly designed, and permissions align with the server's purpose (GPU cluster management). Minor low-severity findings around error handling and logging practices do not materially impact security. Supply chain analysis found 5 known vulnerabilities in dependencies (0 critical, 3 high severity). Package verification found 1 issue.

5 files analyzed · 12 issues found

Security scores are indicators to help you make informed decisions, not guarantees. Always review permissions before connecting any MCP server.

Permissions Required

This plugin requests these system permissions. Most are normal for its category.

HTTP Network Access

Connects to external APIs or services over the internet.

env_vars

Check that this permission is expected for this type of plugin.

File System Read

Reads files on your machine. Normal for tools that analyze or process local data.

File System Write

Writes or modifies files on your machine. Check that this is expected for the tool.

system_info

Check that this permission is expected for this type of plugin.

How to Install

Add this to your MCP configuration file:

{
  "mcpServers": {
    "io-github-aiops-tools-inference-aiops": {
      "args": [
        "inference-aiops"
      ],
      "command": "uvx"
    }
  }
}

Reviews

No reviews yet

Be the first to review this server!

Inference AIops MCP Server - Governed GPU inference ops (vLLM + Ray Serve): latency RCA, | MCP Marketplace