← Browse
model-deployment
HermeticOrmus ★ 0
This plugin guides model deployment across BentoML, TorchServe, Triton Inference Server, and FastAPI, covering ONNX export, dynamic batching, containerization, and Kubernetes rollouts with canary traffic splitting via Istio. It includes patterns for health/readiness probes, Prometheus metrics, HPA configuration, and load testing with locust.
Install
> /plugin marketplace add HermeticOrmus/LibreMLOps-Claude-Code/tree/HEAD/plugins/model-deployment
> /plugin install model-deployment
Source: https://github.com/HermeticOrmus/LibreMLOps-Claude-Code/tree/HEAD/plugins/model-deployment
What it's made of
1 command · 1 agent · 1 skill
- Commands
- 1
- Agents
- 1
- Skills
- 1
- MCP servers
- 0
- Hooks
- 0
What it needs & plugs into
- API keys
- none
- Paid services
- none detected
- External tools
- none
- Talks to
- nothing external detected
Analyzed . Facts extracted from the plugin's files. Prose generated by claude-haiku-4-5-20251001.
Is this your plugin?
Claim it to keep the card accurate and enter the weekly contest. Requires signing in as the GitHub owner (HermeticOrmus).