Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms
GitLab
This broad SRE opening fits IT Support Group readers who combine production operations with software engineering and want to improve reliability at platform scale. GitLab is hiring across intermediate through senior-staff levels for its Infrastructure Platforms teams, with placement based on experience and current team needs.
What you would work on:
- Keep user-facing services and production systems reliable, scalable, and efficient.
- Build infrastructure tooling and automation that replaces manual toil with repeatable workflows.
- Operate and troubleshoot Kubernetes deployments, rollouts, and scaling.
- Ship infrastructure-as-code changes through CI/CD and GitOps practices.
- Improve observability with metrics, logs, alerts, SLOs, incident reviews, and reusable runbooks.
Good fit if:
- You have experience operating reliable production systems with both an operations mindset and software-engineering practice.
- You have built infrastructure tooling such as Terraform modules, Kubernetes operators or controllers, or production automation services.
- You can read and debug code; GitLab notes that most teams use Go and some use Ruby.
- You have hands-on Kubernetes, infrastructure-as-code, and AWS or GCP experience.
- You can contribute at an intermediate through senior-staff scope, from scoped reliability improvements to cross-team technical direction.
Curated from GitLab Careers for IT Support Group readers. This is an external listing; use the apply link for the source listing and latest details.