Site Reliability Engineer
Resolve a service incident with the affected workloads and owners in view.
Design a recoverable reconciliation run
Run the reconciliation kit, then repeat it with the same inputs. Define how the production workflow should handle retries, late settlements, monitoring, and a failed downstream handoff.
Start with the browser demo or download the local kit. For the workspace exercise, you need approved access to the relevant Genedata capabilities and a reviewer for your output.
Practice tasks
0 / 4 completed
Notes and progress are saved on this device.
Move from signals to a clear response.
Resolve a service incident with the affected workloads and owners in view.
Review the service objective, alert, and scope of affected users or workloads.
Correlate operational signals with recent releases, dependency changes, and capacity conditions.
Coordinate the response using the maintained runbook and the agreed escalation path.
Verify recovery against the objective and document the cause, follow-up owner, and prevention work.
Expected handoff
A resolved incident with recovery evidence and owned follow-up work.
Explore your platform capabilitiesMeasure your progress
- Time from an alert to a confirmed diagnosis
- Repeat incidents after a reviewed corrective action
Use AI to assemble a timeline from available evidence; validate causality and keep response decisions with the on-call team.