# Operations Governance

> Service-Ownership, Bereitschaft und Betriebsrisiko als überprüfbare Verantwortungsstruktur führen.

Track: [Monitoring, Troubleshooting & Incident Operations](https://physar.tech/learn/monitoring-incident-operations)  
Kanonische Fassung: https://physar.tech/learn/monitoring-incident-operations/operations-governance  
Stand: 2026-07-29  
Interaktiver Teil: 3 Checks (nur im Browser)

## Verantwortung, die im Incident funktioniert

### Die Betriebsfrage

Eine Alarmregel, ein Runbook und ein SLO sind nur wirksam, wenn Service-Owner, On-call und Eskalation im tatsächlichen Betriebsfenster klar sind.

> **Lernziel:** Du kannst eine fehlende Verantwortungsgrenze als operatives Risiko erkennen und korrigieren.

### Governance-Schleife

Service und kritische Abhängigkeiten erfassen → Owner und Bereitschaft bestätigen → Runbook und Eskalation testen → Risiko und Ausnahmen regelmäßig reviewen

## Quellen

- sre.google/workbook/on-call — https://sre.google/workbook/on-call/
