AWS Certified AI Practitioner
Token-Based Pricing
Learn how generative AI bills, why output tokens cost more than input tokens, and how on-demand, batch, Provisioned Throughput, and prompt caching change the cost of the same workload.
Intermediate 20 minutes 4 Learning Objectives
- Explain how token-based pricing works and why input and output tokens are priced differently
- Estimate the monthly cost of a workload from its token volumes
- Compare on-demand, batch, and Provisioned Throughput and identify which fits a given scenario
- Describe how prompt caching reduces cost and what its limitations are
