Gitlab logo
Est. Monthly
Estimated €7,668 - €12,307
Posted August 7, 2026 · 0 days agoLast seen August 7, 2026Est. expiry September 11, 2026

Site Reliability Engineer

Site Reliability Engineer - Dedicated Hosted Runners
Bangalore, India
Remote · Software Engineering
Full-time · Mid-Level
English
No People Management
How this salary compares
Salary Context: Site Reliability Engineer

Hover or tap a row for full statistics (EUR / month on this chart).

Salary analysis

Compared with the selected benchmark ("Market Average: Mid-Level Level"), this listing's salary midpoint is about 92% lower. The offer sits below the benchmark range (€7,668–€12,307). The offer's range width is broadly in line with the benchmark. This benchmark is based on 1 comparable listings.

Monthly salary comparison for Site Reliability Engineer
MarketLower bound (25th percentile)MedianUpper bound (75th percentile)
Market Average: Site Reliability Engineer€6,010/per month€7,147/per month€10,942/per month
Pay in our data — not quoted in ad (Mid-Level)€639/per month€832/per month€1,026/per month
About the role

We're looking for an Intermediate Site Reliability Engineer to join the Runners Platform team. In this role, you'll build and operate Hosted Runners for GitLab Dedicated, the managed CI/CD compute platform that runs our customers' pipelines inside single-tenant Amazon Web Services (AWS) environments. You'll own infrastructure automation across the full lifecycle: provisioning runner fleets with Terraform, extending the Go tooling and autoscaling stack, and making sure customer continuous integration and continuous delivery (CI/CD) jobs keep running reliably across the platform. What you’ll do - Design, build, and operate AWS infrastructure for Hosted Runners across many single-tenant environments, including Elastic Compute Cloud (EC2), Auto Scaling Groups, Virtual Private Cloud (VPC) networking, subnets, Network Address Translation (NAT), network access control lists, PrivateLink, Identity and Access Management (IAM), and Elastic Container Registry (ECR). - Develop and maintain infrastructure as code using Terraform, contributing to common modules and the deployment tooling that provisions and upgrades runner stacks. - Write Go code for our runner tooling and autoscaling components, including the fleeting instance-lifecycle plugins, zero-downtime deployment command-line interface, and reusable infrastructure toolkits. - Build and improve the GitLab CI/CD pipelines that orchestrate blue/green zero-downtime deployments, automated upgrades, quality assurance validation, and performance testing of runner stacks. - Define and monitor service level objectives for CI job execution, including queue times, job success rates, and fleet saturation, and build the Grafana dashboards, alerts, and runbooks behind them. - Participate in an on-call rotation, handle incidents affecting customer CI/CD workloads, and automate away recurring toil. - Run performance and scale testing that reflects real customer workloads, and tune autoscaling parameters for cost and reliability. - Write documentation and runbooks so the broader team can operate runner stacks consistently. What you’ll bring - Professional experience operating production infrastructure on AWS at scale, including EC2, Auto Scaling Groups, IAM, and VPC networking. - Strong infrastructure-as-code experience with Terraform, including writing and refactoring modules used by other teams. - Proficiency in Go for building and debugging infrastructure tooling, or strong experience in another systems language and willingness to work in Go daily. - Practical knowledge of CI/CD systems and job execution: how pipelines schedule work, how ephemeral build environments get provisioned and torn down, and what makes CI workloads reliable. - Experience with observability practices such as metrics, dashboards, alerting, logging, and service-level-objective-based monitoring, using tools such as Prometheus, Grafana, and OpenSearch. - Experience with on-call rotations and incident management for customer-facing systems. - Strong problem-solving skills, excellent written communication, and comfort working asynchronously across Americas, Europe, Middle East, Africa, and Asia-Pacific time zones. - Direct GitLab Runner experience, familiarity with configuration management such as Ansible, and container tooling such as Docker are a plus.

Job Details

Responsibilities

  • Design, build, and operate AWS infrastructure for Hosted Runners
  • Develop and maintain IaC using Terraform
  • Write Go code for runner tooling and autoscaling components
  • Improve GitLab CI/CD pipelines for zero-downtime deployments
  • Define and monitor SLOs using Grafana and Prometheus
  • Participate in on-call rotation and handle incidents
  • Run performance and scale testing
  • Write documentation and runbooks

Requirements

  • Professional experience operating production infrastructure on AWS at scale (EC2, Auto Scaling Groups, IAM, VPC)
  • Strong infrastructure-as-code experience with Terraform
  • Proficiency in Go or another systems language
  • Practical knowledge of CI/CD systems and job execution
  • Experience with observability tools (Prometheus, Grafana, OpenSearch)
  • Experience with on-call rotations and incident management
  • Strong written communication and comfort with asynchronous work

Skills & Technologies

AWSEC2Auto Scaling GroupsVPCIAMECRTerraformGoGitLab CI/CDPrometheusGrafanaOpenSearchAnsibleDocker
Seen 22 hours agoContent Complete
Gitlab logo
Gitlab · 351 open roles
Top locations: Remote - Global · 309 · Remote - North America · 14 · Remote · 9+7 other locations
View company
Most-hired roles
Backend Engineer
28
Solutions Architect
21
Commercial Account Executive
15
Product Manager
15
Customer Success Manager
13
Role-level mix
Mid-Level (286)Senior (13)Manager (8)

Help us improve JobCrawls — sign in to sync saved jobs across devices, or send feedback anytime.