Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms @GitLab
Devops
Salary usd 126,400 - 3..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 2mths ago

[Hiring] Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms @GitLab

2mths ago - GitLab is hiring a remote Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms. 💸 Salary: usd 126,400 - 314,400 per year 📍Location: USA

Role Description

Site Reliability Engineers keep GitLab's user-facing services and production systems running reliably at scale. They combine software engineering with operational excellence, applying sound engineering principles, automation, and continuous improvement to build, operate, and evolve our production infrastructure.

This is a single application for Site Reliability Engineering opportunities across our Infrastructure Platforms department. Rather than asking you to choose the right team or level upfront, we evaluate your skills holistically and match you to the opportunity that best aligns with your experience and our hiring needs. We hire Site Reliability Engineers from Intermediate through Senior Staff across multiple Infrastructure Platforms teams.

We don't expect every candidate to have experience with every technology in our environment. We're looking for engineers with strong technical fundamentals, a growth mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working.

What you'll do

  • Keep user-facing services and production systems reliable, scalable, and efficient
  • Build automation and tooling that reduces toil and replaces manual work with repeatable, infrastructure-as-code-driven workflows
  • Operate and troubleshoot production systems on Kubernetes, including deployments, rollouts, and scaling
  • Write and maintain infrastructure as code, and ship changes safely through CI/CD and GitOps
  • Participate in on-call, triage alerts, follow and improve runbooks, and escalate appropriately
  • Contribute to the observability stack, using metrics, logs, and SLOs to detect symptoms early rather than just outages
  • Take part in incident response and post-incident reviews, turning learnings into changes in automation and process
  • Document runbooks, architecture decisions, and reviews so your findings become repeatable practices

Qualifications

  • Experience keeping production systems reliable, combining an operations mindset with real software engineering practice
  • Experience building net-new infrastructure tooling and automation, not just configuring existing tools (e.g., Terraform modules, Kubernetes operators or controllers)
  • The ability to read, debug, and reason about code (most teams work in Go; some work in Ruby)
  • Experience with infrastructure as code, and with Kubernetes and its ecosystem, at a depth appropriate to your level
  • Hands-on experience with at least one major cloud provider (GCP or AWS)
  • Familiarity with observability practices, including metrics, logging, alerting, and SLOs or SLIs
  • Comfort participating in on-call and incident response, with a structured approach to troubleshooting under pressure
  • Strong written communication and the ability to operate as a manager-of-one in an async, distributed environment
  • A track record of using automation, and increasingly AI, to reduce toil and improve how you and your team work
  • Alignment with GitLab's values and a commitment to working in accordance with them

Requirements

  • Intermediate: Make meaningful contributions to reliability, automation, and operational efficiency
  • Senior: Drive reliability improvements across multiple projects or services
  • Staff: Shape reliability strategy across teams and services
  • Senior Staff: Set technical direction for reliability across a sub-department

Benefits

  • Flexible Paid Time Off
  • Team Member Resource Groups
  • Equity Compensation & Employee Stock Purchase Plan
  • Growth and Development Fund
  • Parental Leave
Before You Apply
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms @GitLab
Devops
Salary usd 126,400 - 3..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 2mths ago
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
×
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 ★★★★★ from 500+ reviews

⚡ 126,904+ remote jobs, refreshed hourly

🔔 Real-time alerts: Apply first, direct to employer

🛡️ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later