DRHCOPS06-BP01 Develop a process for each alert that you defined for your Outposts and Local Zone workloads
For Outposts, understand who owns different alerts. Instance-level alerts should likely be owned by the account owners, whereas application owners on those resources can stay informed. Loss of network or power should be owned by the team operating the infrastructure.
Desired outcome: Implement well-defined and tested processes. Implement procedures for each alert generated by monitoring systems for Outposts and Local Zone workloads, which provides consistent and effective incident response and remediation.
Benefits of establishing this best practice: Having a structured process for each alert for Outposts and Local Zone workloads enables prompt and coordinated actions, minimizes the risk of overlooking critical issues, and facilitates efficient troubleshooting and resolution.
Level of risk exposed if this best practice is not established: Medium
Implementation guidance
Define runbooks for different scenarios and communication channels that engage the responsible teams to each identified scenario. Set up rack- and instance-level alerts that align with your runbooks.