
Hover or tap a row for full statistics (EUR / month on this chart).
Salary analysis
Compared with the selected benchmark ("In Remote - Europe: Site Reliability Engineer"), this listing's salary midpoint is about 91% lower. The offer sits below the benchmark range (€3,856–€8,670). The listed pay band (€480–€658) is tighter than the benchmark, which suggests lower salary variability. This benchmark is based on 1 comparable listings.
| Market | Lower bound (25th percentile) | Median | Upper bound (75th percentile) |
|---|---|---|---|
| Market Average: Site Reliability Engineer | €5,022/per month | €7,751/per month | €13,799/per month |
| In Remote - Europe: Site Reliability Engineer | €3,856/per month | €6,263/per month | €8,670/per month |
We are looking for a Senior Site Reliability Engineer (SRE) to join the Compute Node team at Nebius AI Cloud. The Compute Node team is responsible for building and operating the cluster scheduler and node-level services that run and manage virtual machines across all cloud regions. This role focuses on Linux systems engineering, virtualization and operational reliability. You will work close to the operating system and hypervisor, shaping how reliability and observability are embedded into the Compute platform. Your responsibilities will include: - Ensure reliability, availability and performance of compute nodes running VMs - Analyze and debug Linux systems across user space and kernel space, understanding capabilities, limitations and trade-offs at each layer - Troubleshoot complex production issues involving CPU, memory, NUMA, cgroups and scheduling - Work hands-on with virtualization and containerization, primarily using QEMU/KVM and Linux-native technologies - Design and evolve observability as a core capability of the node layer: metrics, logs, traces, alerts, SLIs and SLOs - Lead incident response, root-cause analysis, and postmortems, driving long-term reliability improvements - Collaborate closely with platform, kernel/hypervisor, GPU and infrastructure teams to improve system design and operability We expect you to have: - Strong Linux expertise: deep understanding of Linux user space and kernel space; knowledge of kernel subsystems (scheduler, memory management, filesystems, cgroups, namespaces); clear understanding of system boundaries and constraints at different layers - Virtualization experience: hands-on experience with QEMU/KVM; understanding of VM lifecycle, performance characteristics and failure modes - Containerization knowledge: practical experience with containers, namespaces and cgroups; strong understanding of resource isolation and control - Strong debugging skills: ability to reason about complex system failures; structured, hypothesis-driven approach to incident analysis - SRE mindset: clear understanding of the SRE role in system design and operations; experience building and operating observability stacks, not just consuming them; ability to turn system behavior into actionable reliability signals Nice to Have / Optional: - Experience with Kubernetes internals or node-level components - Hands-on experience with low-level Linux debugging tools (e.g. perf, eBPF, ftrace, strace, kernel crash dumps) - Familiarity with large-scale compute or bare-metal platforms - Contributions to open-source infrastructure or system software - Experience debugging hardware and driver-level issues, including GPUs, NVLink, InfiniBand Benefits & Perks: - Competitive compensation - Career growth and learning opportunities - Flexibility and ownership - Collaborative and innovative culture - Opportunity to work on impactful AI projects - International environment and talented teams What it’s like to work at Nebius: - Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.
Job Details
Responsibilities
- Ensure reliability, availability and performance of compute nodes running VMs
- Analyze and debug Linux systems across user space and kernel space
- Troubleshoot complex production issues involving CPU, memory, NUMA, cgroups and scheduling
- Work hands-on with virtualization and containerization, primarily using QEMU/KVM and Linux-native technologies
- Design and evolve observability as a core capability of the node layer: metrics, logs, traces, alerts, SLIs and SLOs
- Lead incident response, root-cause analysis, and postmortems, driving long-term reliability improvements
- Collaborate closely with platform, kernel/hypervisor, GPU and infrastructure teams to improve system design and operability
Requirements
- Strong Linux expertise: deep understanding of Linux user space and kernel space
- Knowledge of kernel subsystems (scheduler, memory management, filesystems, cgroups, namespaces)
- Virtualization experience: hands-on experience with QEMU/KVM
- Containerization knowledge: practical experience with containers, namespaces and cgroups
- Strong debugging skills: ability to reason about complex system failures
- SRE mindset: observability stacks experience, ability to turn system behavior into signals
Skills & Technologies
Benefits & Perks
Recruitment Process
- 1Submit application
- 2Phone screen
- 3Technical interview
- 4Onsite interview

| Location | Active listings |
|---|---|
| Remote - Global | 423 |
| Remote - Europe | 119 |
| Amsterdam, Netherlands | 56 |
| London, United Kingdom | 22 |
| Remote - United States | 20 |
| Remote - Finland | 18 |
| Berlin, Germany | 15 |
| Mäntsälä, Finland | 12 |
| Helsinki, Finland | 10 |
| Lappeenranta, Finland | 9 |
| Prague, Czech Republic | 6 |
| Amsterdam | 5 |
| Israel | 5 |
| United Kingdom | 4 |
| Canada | 4 |
| Remote | 3 |
| Tel Aviv, Israel | 3 |
| Singapore | 3 |
| Remote - United Kingdom | 2 |
| Abu Dhabi | 2 |
| Dubai | 2 |
| Paris, France | 2 |
| Remote - France | 2 |
| New York City, United States | 2 |
| Austin, United States | 2 |
| France, Paris | 2 |
| Remote - Germany | 2 |
| London | 2 |
| Remote - Netherlands | 2 |
| Philadelphia, United States | 1 |
| Béthune, France | 1 |
| Abu Dhabi, Dubai | 1 |
| Singapore, Singapore | 1 |
| California, United States | 1 |
| Abu Dhabi, United Arab Emirates | 1 |
| Remote - EU | 1 |
| Munich, Germany | 1 |
| Minnesota, United States | 1 |
| Alabama, US | 1 |
| Prague, Czechia | 1 |
| East London, United Kingdom | 1 |
| Canada, Remote - United States | 1 |
| Berlin | 1 |
| Dallas, United States | 1 |
| London, UK | 1 |
| Oklahoma, United States | 1 |
| New Jersey, US | 1 |
| Austin, Texas | 1 |
| Kansas City, United States | 1 |
| San Francisco Bay Area, United States | 1 |
| Czechia | 1 |
| Finland | 1 |
| Remote - Sweden | 1 |
| UK | 1 |
| Prague | 1 |
| New Jersey, United States | 1 |
| Remote - Czech Republic | 1 |
| Netherlands | 1 |
| Paris | 1 |
| Béthune, Pas-de-Calais, France | 1 |
| Role type | Active listings |
|---|---|
| Backend Engineer | 331 |
| Software Engineer | 82 |
| Account Executive | 58 |
| Technical Project Manager | 6 |
| Site Reliability Engineer | 4 |
| Technical Product Manager | 4 |
| System Engineer | 4 |
| Sales Representative | 4 |
| Product Manager | 3 |
| Data Center Operations Technician | 3 |
| Data Center Technician | 3 |
| ML Engineer | 3 |
| Technical Program Manager | 3 |
| Delivery Manager | 2 |
| Applied AI Researcher | 2 |
| Hypervisor Engineer | 2 |
| Product Designer | 2 |
| IT Technician | 2 |
| Backend engineers, Frontend engineers, Site reliability engineers | 2 |
| Open Positions at Nebius | 2 |
| VP of Strategic Sales | 1 |
| Generalist | 1 |
| Offensive Security Lead | 1 |
| Head of Channel Marketing | 1 |
| Principal | 1 |
| Applied AI Solutions Engineer | 1 |
| Field Technical Lead | 1 |
| Compensation Analyst | 1 |
| Solutions Architecture Leader | 1 |
| Application Security Engineer | 1 |
| Human Resources Specialist | 1 |
| Solutions Architect | 1 |
| Mechanical Data Center Technician | 1 |
| Backend Developer | 1 |
| Senior Research Scientist | 1 |
| Cloud Solution Architect | 1 |
| Data Center IT Technician | 1 |
| Manager, ML Solutions Architecture | 1 |
| ML Solutions Architect | 1 |
| Network Planning Project Manager | 1 |
| Data Center Logistics Specialist | 1 |
| Technical Due Diligence Manager | 1 |
| Data Center Operations Manager | 1 |
| Senior Site Reliability Engineer | 1 |
| IT Support Manager | 1 |
| Technical Support Engineer | 1 |
| Data Engineer | 1 |
| Mechanical Engineer | 1 |
| Data Center IT Manager | 1 |
| Security Product Manager | 1 |
| Data Center Electrical Lead | 1 |
| GTM Recruiting Manager | 1 |
| Security Solutions Engineer | 1 |
| Physical Security Systems Technician | 1 |
| Senior HPC Engineer | 1 |
| Educational Content Author | 1 |
| Customer Engineer | 1 |
| Pricing Director | 1 |
| MEP Engineer | 1 |
| Instructional Designer | 1 |
| Partner Solutions Architect | 1 |
| Senior Software Developer | 1 |
| Forward Deployment Engineer | 1 |
| Financial Controller | 1 |
| Internal Control Business Partner | 1 |
| Data Center Facilities Manager | 1 |
| Group Product Manager | 1 |
| Structured Cabling Design Engineer | 1 |
| Site Selection & Colocation Manager | 1 |
| Product Growth Analytics Lead | 1 |
| IT Risk and Control Manager | 1 |
| Vulnerability Operations Center Lead | 1 |
| Electrical Engineer | 1 |
| Senior System Engineer | 1 |
| Machine Learning Engineer | 1 |
| ML Infrastructure Engineer | 1 |
| Data Scientist | 1 |
| Mechanical Design Engineer | 1 |
| Applied ML Engineer | 1 |
| Deal Initiation and Activation Manager | 1 |
| Operations Specialist | 1 |
| AI/ML Specialist Solutions Architect | 1 |
| HPC Engineer | 1 |
| Accountant | 1 |
| VP of Developer Relations & Community | 1 |
| Solutions Partner | 1 |
| Senior Support Engineer | 1 |
| Network Engineer | 1 |
| Project Development Manager | 1 |
| Backend Engineers | 1 |
| Role level | Active listings |
|---|---|
| Mid-Level | 410 |
| Senior | 63 |
| Manager | 12 |
| Executive | 2 |
| Director | 1 |
Related Opportunities
Discover more opportunities that match your interests and skills