Create your own

Site Reliability Engineering Interview Preparation

Service Levels and Error-Budget Decisions
Datadog Metrics and Operational Dashboards
Logs, Traces, APM, and Real-User Signals
Actionable Alert Engineering
On-Call Readiness and Incident Command
Blameless Learning After Incidents
Production Troubleshooting and Performance Tuning
Resilient Software and Capacity Planning
Toil Reduction, Terraform, and Safe Delivery
Independent Reliability Initiative Capstone
SRE Interview and Evidence Preparation