Cost, Latency, Caching, and Rate Limits
By the end of this module you can measure and control cost and latency, and use caching and rate limits without silently serving a stale answer.
Units
- Unit 08.00: Measuring cost per request honestly
- Unit 08.01: Where the latency actually goes
- Unit 08.02: Caching without serving a stale answer
- Unit 08.03: Rate limits and graceful degradation
- Unit 08.04: The optimisation that changed the output
