# Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms

**Company**: GitLab
**Location**: Remote
**Work arrangement**: remote
**Job type**: full-time
**Salary**: $126,400-$314,400 USD
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/gitlab/jobs/8623389002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_19dfe537-35d

## Description

GitLab is seeking Site Reliability Engineers to join their Infrastructure Platforms department. As a Site Reliability Engineer, you will keep GitLab's user-facing services and production systems running reliably at scale. You will combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve production infrastructure.

Responsibilities:

- Keep user-facing services and production systems reliable, scalable, and efficient

- Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows

- Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling

- Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps

- Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately

- Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages

- Take part in incident response and post-incident reviews, turning learnings into changes in automation and process

- Document runbooks, architecture decisions, and reviews so your findings become repeatable practices

Requirements:

- Experience keeping production systems reliable, combining an operations mindset with real software engineering practice

- Experience building net-new infrastructure tooling and automation, not just configuring existing tools

- The ability to read, debug, and reason about code

- Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level

- Hands-on experience with at least one major cloud provider (GCP or AWS)

- Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs, and using data to inform operational decisions

- Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure

- Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment

- A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work

- Alignment with GitLab's values and a commitment to working in accordance with them

The base salary range for this role is $126,400-$314,400 USD.

GitLab supports full-time employees with benefits to support their health, finances, and well-being, flexible paid time off, team member resource groups, equity compensation and employee stock purchase plan, growth and development fund, and parental leave.

## Skills

### Required
- Kubernetes
- infrastructure as code
- CI/CD
- GitOps
- observability
- metrics
- logs
- SLOs
- SLIs
- cloud provider
- automation
- software engineering

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/gitlab/jobs/8623389002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
