AWS Certified AI Practitioner

Token-Based Pricing

Learn how generative AI bills, why output tokens cost more than input tokens, and how on-demand, batch, Provisioned Throughput, and prompt caching change the cost of the same workload.

Intermediate 20 minutes 4 Learning Objectives
  1. Explain how token-based pricing works and why input and output tokens are priced differently
  2. Estimate the monthly cost of a workload from its token volumes
  3. Compare on-demand, batch, and Provisioned Throughput and identify which fits a given scenario
  4. Describe how prompt caching reduces cost and what its limitations are