AWS Certified AI Practitioner
Types of Inferencing
Once a model is trained, inference is how it serves predictions. Compare real-time, batch, asynchronous, and serverless inference and match each to latency, payload, and traffic needs.
Beginner 15 minutes 4 Learning Objectives
- Define inference and distinguish it from training
- Compare real-time, batch, asynchronous, and serverless inference
- Match an inference type to latency, payload, and traffic requirements
- Map the inference types to Amazon SageMaker options
