45 karma · joined February 14, 2019
Cortex is what they are referring to as a downstream tool for real-time serving through a REST API. In other words, MLflow helps with model management and packaging, whereas Cortex is a platform for running real-time inference at scale. We are working on supporting more model packaging formats and I think it's a good idea to support the MLflow format as well.
- Deployments are defined with declarative configuration and no custom Docker images are required (although you can use your own if you want)
- You have full access to the instances, autoscaling groups, security groups, etc
- Less tied to AWS (GCP support is in the works)
- We are working on higher level features like prediction monitoring, alerting, and model retraining
- It's open source and free vs SageMaker's ~40% markup