ADVPERF01-BP04 Evaluate AI/ML-based architecture for optimization (like contextual advertising or scaling algorithms on event context)
Use AWS services to implement a low latency, high throughput inference and MLOps framework.
Implementation guidance
-
Implement low-latency, high-throughput model inference using Amazon ECS
, Amazon EKS , and Amazon SageMaker AI . -
Implement an ML pipeline using Amazon SageMaker AI to build, train, and deploy machine learning models. Additionally, use Sage Maker for predictive scaling of compute based on learning from past event data.
Resources
Related documentation:
Related videos: