Nebius B.V. logo
Est. Monthly
Estimated €5,760 - €7,901
Posted August 28, 2026 · 0 days agoLast seen August 28, 2026Est. expiry October 2, 2026

Site Reliability Engineer

Senior Site Reliability Engineer (SRE
How this salary compares
Salary Context: Site Reliability Engineer

Hover or tap a row for full statistics (EUR / month on this chart).

Salary analysis

Compared with the selected benchmark ("In Remote - Europe: Site Reliability Engineer"), this listing's salary midpoint is about 91% lower. The offer sits below the benchmark range (€3,856–€8,670). The listed pay band (€480–€658) is tighter than the benchmark, which suggests lower salary variability. This benchmark is based on 1 comparable listings.

Monthly salary comparison for Site Reliability Engineer
MarketLower bound (25th percentile)MedianUpper bound (75th percentile)
Market Average: Site Reliability Engineer€5,022/per month€7,751/per month€13,799/per month
In Remote - Europe: Site Reliability Engineer€3,856/per month€6,263/per month€8,670/per month
About the role

We are looking for a Senior Site Reliability Engineer (SRE) to join the Compute Node team at Nebius AI Cloud. The Compute Node team is responsible for building and operating the cluster scheduler and node-level services that run and manage virtual machines across all cloud regions. This role focuses on Linux systems engineering, virtualization and operational reliability. You will work close to the operating system and hypervisor, shaping how reliability and observability are embedded into the Compute platform. Your responsibilities will include: - Ensure reliability, availability and performance of compute nodes running VMs - Analyze and debug Linux systems across user space and kernel space, understanding capabilities, limitations and trade-offs at each layer - Troubleshoot complex production issues involving CPU, memory, NUMA, cgroups and scheduling - Work hands-on with virtualization and containerization, primarily using QEMU/KVM and Linux-native technologies - Design and evolve observability as a core capability of the node layer: metrics, logs, traces, alerts, SLIs and SLOs - Lead incident response, root-cause analysis, and postmortems, driving long-term reliability improvements - Collaborate closely with platform, kernel/hypervisor, GPU and infrastructure teams to improve system design and operability We expect you to have: - Strong Linux expertise: deep understanding of Linux user space and kernel space; knowledge of kernel subsystems (scheduler, memory management, filesystems, cgroups, namespaces); clear understanding of system boundaries and constraints at different layers - Virtualization experience: hands-on experience with QEMU/KVM; understanding of VM lifecycle, performance characteristics and failure modes - Containerization knowledge: practical experience with containers, namespaces and cgroups; strong understanding of resource isolation and control - Strong debugging skills: ability to reason about complex system failures; structured, hypothesis-driven approach to incident analysis - SRE mindset: clear understanding of the SRE role in system design and operations; experience building and operating observability stacks, not just consuming them; ability to turn system behavior into actionable reliability signals Nice to Have / Optional: - Experience with Kubernetes internals or node-level components - Hands-on experience with low-level Linux debugging tools (e.g. perf, eBPF, ftrace, strace, kernel crash dumps) - Familiarity with large-scale compute or bare-metal platforms - Contributions to open-source infrastructure or system software - Experience debugging hardware and driver-level issues, including GPUs, NVLink, InfiniBand Benefits & Perks: - Competitive compensation - Career growth and learning opportunities - Flexibility and ownership - Collaborative and innovative culture - Opportunity to work on impactful AI projects - International environment and talented teams What it’s like to work at Nebius: - Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

Job Details

Responsibilities

  • Ensure reliability, availability and performance of compute nodes running VMs
  • Analyze and debug Linux systems across user space and kernel space
  • Troubleshoot complex production issues involving CPU, memory, NUMA, cgroups and scheduling
  • Work hands-on with virtualization and containerization, primarily using QEMU/KVM and Linux-native technologies
  • Design and evolve observability as a core capability of the node layer: metrics, logs, traces, alerts, SLIs and SLOs
  • Lead incident response, root-cause analysis, and postmortems, driving long-term reliability improvements
  • Collaborate closely with platform, kernel/hypervisor, GPU and infrastructure teams to improve system design and operability

Requirements

  • Strong Linux expertise: deep understanding of Linux user space and kernel space
  • Knowledge of kernel subsystems (scheduler, memory management, filesystems, cgroups, namespaces)
  • Virtualization experience: hands-on experience with QEMU/KVM
  • Containerization knowledge: practical experience with containers, namespaces and cgroups
  • Strong debugging skills: ability to reason about complex system failures
  • SRE mindset: observability stacks experience, ability to turn system behavior into signals

Skills & Technologies

LinuxQEMU/KVMContainersNamespacesCGroupseBPFperfftracekernel crash dumps

Recruitment Process

  1. 1
    Submit application
  2. 2
    Phone screen
  3. 3
    Technical interview
  4. 4
    Onsite interview
Seen 16 hours agoPartial Schema
Nebius B.V. logo
Nebius B.V. · 790 open roles
Top locations: Remote - Global · 423 · Remote - Europe · 119 · Amsterdam, Netherlands · 56+57 other locations
View company
Current open roles at Nebius B.V. on JobCrawls
LocationActive listings
Remote - Global423
Remote - Europe119
Amsterdam, Netherlands56
London, United Kingdom22
Remote - United States20
Remote - Finland18
Berlin, Germany15
Mäntsälä, Finland12
Helsinki, Finland10
Lappeenranta, Finland9
Prague, Czech Republic6
Amsterdam5
Israel5
United Kingdom4
Canada4
Remote3
Tel Aviv, Israel3
Singapore3
Remote - United Kingdom2
Abu Dhabi2
Dubai2
Paris, France2
Remote - France2
New York City, United States2
Austin, United States2
France, Paris2
Remote - Germany2
London2
Remote - Netherlands2
Philadelphia, United States1
Béthune, France1
Abu Dhabi, Dubai1
Singapore, Singapore1
California, United States1
Abu Dhabi, United Arab Emirates1
Remote - EU1
Munich, Germany1
Minnesota, United States1
Alabama, US1
Prague, Czechia1
East London, United Kingdom1
Canada, Remote - United States1
Berlin1
Dallas, United States1
London, UK1
Oklahoma, United States1
New Jersey, US1
Austin, Texas1
Kansas City, United States1
San Francisco Bay Area, United States1
Czechia1
Finland1
Remote - Sweden1
UK1
Prague1
New Jersey, United States1
Remote - Czech Republic1
Netherlands1
Paris1
Béthune, Pas-de-Calais, France1
Current role mix at Nebius B.V. on JobCrawls
Role typeActive listings
Backend Engineer331
Software Engineer82
Account Executive58
Technical Project Manager6
Site Reliability Engineer4
Technical Product Manager4
System Engineer4
Sales Representative4
Product Manager3
Data Center Operations Technician3
Data Center Technician3
ML Engineer3
Technical Program Manager3
Delivery Manager2
Applied AI Researcher2
Hypervisor Engineer2
Product Designer2
IT Technician2
Backend engineers, Frontend engineers, Site reliability engineers2
Open Positions at Nebius2
VP of Strategic Sales1
Generalist1
Offensive Security Lead1
Head of Channel Marketing1
Principal1
Applied AI Solutions Engineer1
Field Technical Lead1
Compensation Analyst1
Solutions Architecture Leader1
Application Security Engineer1
Human Resources Specialist1
Solutions Architect1
Mechanical Data Center Technician1
Backend Developer1
Senior Research Scientist1
Cloud Solution Architect1
Data Center IT Technician1
Manager, ML Solutions Architecture1
ML Solutions Architect1
Network Planning Project Manager1
Data Center Logistics Specialist1
Technical Due Diligence Manager1
Data Center Operations Manager1
Senior Site Reliability Engineer1
IT Support Manager1
Technical Support Engineer1
Data Engineer1
Mechanical Engineer1
Data Center IT Manager1
Security Product Manager1
Data Center Electrical Lead1
GTM Recruiting Manager1
Security Solutions Engineer1
Physical Security Systems Technician1
Senior HPC Engineer1
Educational Content Author1
Customer Engineer1
Pricing Director1
MEP Engineer1
Instructional Designer1
Partner Solutions Architect1
Senior Software Developer1
Forward Deployment Engineer1
Financial Controller1
Internal Control Business Partner1
Data Center Facilities Manager1
Group Product Manager1
Structured Cabling Design Engineer1
Site Selection & Colocation Manager1
Product Growth Analytics Lead1
IT Risk and Control Manager1
Vulnerability Operations Center Lead1
Electrical Engineer1
Senior System Engineer1
Machine Learning Engineer1
ML Infrastructure Engineer1
Data Scientist1
Mechanical Design Engineer1
Applied ML Engineer1
Deal Initiation and Activation Manager1
Operations Specialist1
AI/ML Specialist Solutions Architect1
HPC Engineer1
Accountant1
VP of Developer Relations & Community1
Solutions Partner1
Senior Support Engineer1
Network Engineer1
Project Development Manager1
Backend Engineers1
Current role-level mix at Nebius B.V. on JobCrawls
Role levelActive listings
Mid-Level410
Senior63
Manager12
Executive2
Director1

Help us improve JobCrawls — sign in to sync saved jobs across devices, or send feedback anytime.