Inference AIops

本地

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

⭐ 0其他仓库 ↗

安装配置

{
  "mcpServers": {
    "inference-aiops": {
      "command": "npx",
      "args": [
        "inference-aiops"
      ]
    }
  }
}

相关服务器