IN

Inference AIops MCP

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

not scanned Local · stdio No auth needed tools not listed v0.9.0
No reviews yet by AIops-toolsPyPI 6,420 Runs on your machine, nothing hosted
p95 latency
call success
local
runs on your machine
calls last 7d

What it does

Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

Quickstart

# 1 — run it from where its publisher ships it uvx inference-aiops # 2 — ask your agent something > Governed GPU inference ops (vLLM + Ray Serve): latency RCA, scaling, drain, 39 tools.

Inference AIops is free: there is no plan to choose, no cap to set and nothing that can bill you.

Publishernot claimed
AT
AIops-tools25 servers
8 categories
All 25 servers

Collected from a public index. Nobody has claimed this account, so nothing here was written by its author — claim it if it is yours.

More in this category

See also

See all