Senior Site Reliability Engineer
Veeam Software
This senior SRE role fits IT Support Group readers who want to own reliability for regulated cloud infrastructure. You would help build Veeam’s reliability practice for its Government and Sovereign Cloud environment, working across the full platform stack where resilience, documentation, compliance, and end-to-end operational ownership matter.
**What you would work on:**
- Map platform dependencies and risks, then create runbooks, architecture documentation, and operational guides.
- Design highly available Azure and Azure Government infrastructure and define SLIs, SLOs, and error budgets.
- Lead incident response and blameless postmortems, close observability gaps, and automate repetitive operations.
- Improve IaC, CI/CD, deployment validation, configuration management, and on-call practices in restricted environments.
**Good fit if:**
- You have 7+ years in software engineering, including 3+ years in SRE, platform engineering, or a similar multi-service environment.
- You have operated Government or Sovereign Cloud and understand regulated frameworks such as FedRAMP, CMMC, PCI-DSS, SOX, HIPAA, or HITRUST.
- You know Azure cloud services, Terraform or comparable IaC, Kubernetes, CI/CD or GitOps, and observability tooling.
- You can investigate complex systems independently, program in a language such as Go, Java, C#, or TypeScript, and own problems across engineering, security, compliance, and operations.
Curated from Himalayas for IT Support Group readers. This is an external listing; use the apply link for the source listing and latest details.