Provide oversight of incident, problem, and availability risk trends, driving accountability for remediation actions, root cause quality, service stability improvements, and control compliance across engineering and platform teams. Drive the design and implementation of enterprise scorecards, governance reporting, KPIs, KRIs, and executive dashboards that provide actionable insights into production resilience, incident performance, availability trends, and problem management effectiveness.