Ml Model Serving

Use this skill when deploying models for inference: TorchServe, BentoML, Ray Serve, KServe, Seldon Core, model deployment, A/B testing, autoscaling, canary deployment, model versioning, inference optimization. This skill enforces: framework selection based on model type, deployment strategy documentation, autoscaling configuration, model versioning scheme, rollback procedure, inference optimization. Do NOT use for: model training, feature store serving, pipeline orchestration, prompt engineering.

j4flmao Updated

File contents

j4flmao/agent-skills/tree/main/skills/ml/model-serving commit ff21866a25

Frequently asked questions

npx skillmds@latest add j4flmao/ml-model-serving