
Infrastructure Operations Engineer
Hover or tap a row for full statistics (EUR / month on this chart).
Salary analysis
Compared with the selected benchmark ("Company in London, United Kingdom"), this listing's salary midpoint is about 13% lower. The offer still falls within the benchmark range (€11,936–€22,426). The listed pay band (€13,333–€16,667) is tighter than the benchmark, which suggests lower salary variability. This benchmark is based on 1 comparable listings.
| Market | Lower bound (25th percentile) | Median | Upper bound (75th percentile) |
|---|---|---|---|
| All roles in London, United Kingdom | €4,030/per month | €9,310/per month | €15,737/per month |
| Company in London, United Kingdom | €11,936/per month | €17,181/per month | €22,426/per month |
Lightning AI is seeking an experienced Infrastructure Operations Engineers to help scale and operate our next-generation AI infrastructure platform. Our InfraOps team sits at the center of reliability, automation, and operational scale for GPU infrastructure. This team owns break/fix operations, incident response, customer provisioning, observability, and the automation systems that keep complex infrastructure running efficiently. In this role, you’ll work hands-on with large-scale GPU environments, Linux systems, bare metal infrastructure, provisioning workflows, and platform reliability. You’ll partner closely with Infrastructure Engineering, Network Operations, and Software Platform teams to troubleshoot issues, improve operational efficiency, and build automation that reduces manual toil over time. This role is based in one of our hubs (NYC, SF, Seattle, or London), with a minimum of 2 in-office days per week and occasional team and company offsites. We are not able to provide visa sponsorship for this position at this time.
Job Details
Responsibilities
- Design, build, and roll out new platforms and patterns to minimize incidents and enable features
- Deploy updates to support internal and end customer use cases
- Collaborate with cross-functional teams across Infrastructure Engineering, Network Operations, and Software Platform
- Participate in on-call rotation and incident response
Requirements
- 8+ years working with Linux as a server / hosting platform
- 5+ years experience with AWS
- 2+ years experience with Kubernetes and strong container fundamentals
- 2+ years experience with Terraform and Ansible
- 2+ years with network attached storage management (via NFS, ceph, or other protocols)
- Experience with monitoring systems (Prometheus, ELK stack)
- Familiarity with the gitops workflow
- Software development experience using Python, Go, bash, or other languages for automation & APIs
- Strong networking fundamentals
Skills & Technologies

| Location | Active listings |
|---|---|
| London, United Kingdom | 1 |
| New York, United States | 1 |
| San Francisco, United States | 1 |
| Seattle, United States | 1 |
| Remote - Global | 1 |
| Role type | Active listings |
|---|---|
| Research Engineer | 1 |
| Role level | Active listings |
|---|---|
| Mid-Level | 1 |
Related Opportunities
Discover more opportunities that match your interests and skills