
Hover or tap a row for full statistics (EUR / month on this chart).
Salary analysis
Compared with the selected benchmark ("All roles in Singapore, Singapore"), this listing's salary midpoint is about 92% lower. The offer sits below the benchmark range (€8,681–€13,311). Range-width comparison is limited because one of the salary bands is incomplete. This benchmark is based on 2 comparable listings.
| Market | Lower bound (25th percentile) | Median | Upper bound (75th percentile) |
|---|---|---|---|
| All roles in Singapore, Singapore | €8,681/per month | €10,996/per month | €13,311/per month |
| Pay in our data — not quoted in ad (Mid-Level) | €916/per month | €916/per month | €916/per month |
We are seeking a GPU Cluster Architect to drive the design of our next-generation AI infrastructure. In this high-impact, hands-on role, you will make end-to-end architectural decisions across compute, networking, and storage — ensuring our platforms can meet the massive scale, performance, and reliability requirements of modern AI workloads. This is a high-impact, hands-on architecture role where you’ll define how tens of thousands of GPUs are interconnected, cooled down, powered, and optimized across multiple data center sites. Your responsibilities will include: Cluster Design: Architect scalable GPU cluster topologies including compute nodes, interconnect (InfiniBand, Ethernet), storage, and control planes. Performance Modeling: Analyze AI/ML workloads (e.g. LLM training, inference) to inform design tradeoffs across latency, bandwidth, and GPU density. Network Architecture: Align with network architect relevant design and validate low-latency, high-throughput interconnects (e.g., InfiniBand HDR/NDR, RoCEv2) at POD and DC scale. Storage Integration: Work with storage teams to optimize performance for training datasets, checkpointing, and others. Reliability & Monitoring: Understand and analyze signal from monitoring systems to the detect flows in design Collaboration: Partner with site reliability, networking, storage, and DC engineering teams to operationalize and scale your architecture. We expect you to have: 5+ years of experience designing clusters. Deep understanding of modern GPU architecture (NVIDIA, AMD, etc.). Experience with HPC interconnects (InfiniBand & RoCE). Solid background in systems architecture, networking, and hardware reliability. Experience in scripting for automation and telemetry pipelines (Python, Go, etc.)
Job Details
Responsibilities
- Architect scalable GPU cluster topologies including compute nodes, interconnects, storage, and control planes
- Analyze AI/ML workloads to inform design tradeoffs across latency, bandwidth, and GPU density
- Validate low-latency, high-throughput interconnects at POD and DC scale
- Optimize performance for training datasets and checkpointing with storage teams
- Analyze monitoring signals to detect design flaws
- Partner with SRE, networking, storage, and DC engineering teams to operationalize architecture
Requirements
- 5+ years of experience designing clusters
- Deep understanding of modern GPU architecture (NVIDIA, AMD, etc.)
- Experience with HPC interconnects (InfiniBand & RoCE)
- Solid background in systems architecture, networking, and hardware reliability
- Experience in scripting for automation and telemetry pipelines (Python, Go, etc.)
Skills & Technologies

Related Opportunities
Discover more opportunities that match your interests and skills