Skip to content

Model Deployment & Serving

Coming soon.

  • Batch vs. real-time serving — which one a given use case actually needs
  • Canary and shadow rollout for a new model version before it takes full traffic
  • Rollback: reverting a bad deployment fast, with minimal blast radius
  • The ML-specific layer on top of what Infrastructure already covers for deployment in general

Part of MLOps → AI Engineering.