Production operations
Provision, deploy, observe, and recover a running application.
Provision, deploy, observe, and recover the application
The application must keep a clear contract during a release, a slow dependency, or a growing queue. You will connect it to infrastructure, release compatible changes, record useful evidence, and recover from failures.
Cloud labs are identified explicitly and include their resource setup and cleanup. A local protocol model is useful evidence, but it is not a deployed fleet or a production capacity measurement.
Parts group related chapters. Each lesson has a chapter.lesson address, such as 4.07. Open a title below, or use Next to follow the reading sequence. Within a lesson, On this page lists its sections.
- CH 12
AWS infrastructure
Map application mechanisms to AWS resources, permissions, and limits.
- CH 13
Delivery and controlled rollouts
Build once, preserve compatibility, and release with stop conditions.
- CH 14
Production observability
Use logs, metrics, and traces to explain production behavior.
- CH 15
Reliability and incident recovery
Set reliability objectives, bound overload, and recover from incidents.