Nebius B.V. logo
Monthly
€15,047 - €18,881
Posted August 29, 2026 · 0 days agoLast seen August 29, 2026Est. expiry October 3, 2026

Principal ML Solutions Architect

How this salary compares
Salary Context: Principal ML Solutions Architect

Hover or tap a row for full statistics (EUR / month on this chart).

Salary analysis

Compared with the selected benchmark ("Company in Remote - United States"), this listing's salary midpoint is about 88% higher. The offer sits above the benchmark range (€1,667–€19,150). The listed pay band (€17,333–€21,750) is tighter than the benchmark, which suggests lower salary variability. This benchmark is based on 5 comparable listings.

Monthly salary comparison for Principal ML Solutions Architect
MarketLower bound (25th percentile)MedianUpper bound (75th percentile)
All roles in Remote - United States€1,667/per month€9,586/per month€22,616/per month
Company in Remote - United States€1,667/per month€1,910/per month€19,150/per month
About the role

The role sits within Nebius Token Factory, our serverless platform for running and customizing open-source LLMs in production. Token Factory enables serverless inference and fine-tuning (LoRA, full FT, RFT) with internal optimizations such as speculative decoding, quantization, cache-aware routing and dedicated endpoints. You will be the most senior technical authority for customers leveraging Token Factory’s serverless inference and fine-tuning platforms. Beyond designing and implementing optimized inference and fine-tuning workflows, you will set technical direction across our largest and most strategic accounts, own the hardest performance and quality problems end to end, mentor other Solutions Architects, and shape the platform roadmap with backend, product, and research teams. You may work remotely from the United States. Your responsibilities will include: Own the most complex, highest-stakes customer engagements from architecture through production across multiple modalities, driving measurable business value. Optimize LLM inference at the framework and hardware level and codify the resulting best practices into reusable playbooks. Lead supervised and reinforcement fine-tuning efforts to maximize model quality. Design and implement production-ready LLM solutions using Token Factory's inference services. Provide deep technical expertise in prompt engineering, RAG architectures, model selection, and cost/performance trade-offs at scale. Partner closely with product, engineering and research to surface customer needs, prototype platform features, and directly influence the roadmap. Guide customers from PoC to production with a focus on performance, reliability, and cost efficiency — and define the standards by which the team does so. Mentor Senior and mid-level Solutions Architects; raise the technical bar of the team through review, enablement, and knowledge sharing. Represent Token Factory externally through talks, blog posts, and conferences. We expect you to have: 8+ years of experience in ML/AI systems, with at least 4 years focused on LLMs and generative AI. Demonstrated technical leadership: owning ambiguous, high-impact problems end to end and influencing decisions across teams and customers. Expert knowledge of the LLM ecosystem: model architectures, fine-tuning approaches, and inference internals. Deep, hands-on command of inference optimization: quantization, KV-cache management, batching, routing, etc. Hands-on experience with: Running LLMs in production at scale, LLM fine-tuning including SFT/LoRA, LLM evaluation and deployment of LLM-powered applications. Strong Python programming skills and excellent communication. Nice-to-have: Contributions to OSS inference/ML projects, published research, multimodal AI, DevOps tooling, internal tooling for ML workflows. Preferred tech stack includes Python, vLLM, TensorRT-LLM, SGLang, Transformers, OpenAI/Anthropic SDKs, Kubernetes, Docker, cloud platforms. Key Employee Benefits and Pay Transparency outlined with salary: 208k-261k USD base.

Job Details

Responsibilities

  • Own the most complex, high-stakes customer engagements from architecture through production across multiple modalities
  • Optimize LLM inference at framework and hardware level and codify best practices into reusable playbooks
  • Lead supervised and reinforcement fine-tuning to maximize model quality
  • Design and implement production-ready LLM solutions using Token Factory's inference services
  • Provide deep technical expertise in prompt engineering, RAG architectures, model selection, and cost/performance trade-offs at scale
  • Partner with product, engineering and research to surface customer needs and influence roadmap
  • Guide customers from PoC to production focusing on performance and cost efficiency
  • Mentor Senior and mid-level Solutions Architects; raise technical bar through reviews and enablement
  • Represent Token Factory externally through talks, blogs, and conferences

Requirements

  • 8+ years of experience in ML/AI systems, with at least 4 years focused on LLMs and generative AI
  • Demonstrated technical leadership: owning ambiguous, high-impact problems end to end and influencing decisions across teams and customers
  • Expert knowledge of the LLM ecosystem: model architectures, fine-tuning approaches, and inference internals
  • Deep hands-on command of inference optimization: quantization, KV-cache management, batching, routing
  • Experience with running LLMs in production at scale; LLM fine-tuning including SFT/LoRA; LLM evaluation and deployment of LLM-powered apps

Skills & Technologies

PythonvLLMTensorRT-LLMSGLangTransformersOpenAI/Anthropic SDKsKubernetesDockerAWS SageMaker/BedrockVertex AIAzure ML

Recruitment Process

  1. 1
    Submit application
  2. 2
    Screening call
  3. 3
    Technical interview
  4. 4
    Offer
Seen 8 hours agoPartial Schema
Nebius B.V. logo
Nebius B.V. · 772 open roles
Top locations: Remote - Global · 398 · Remote - Europe · 126 · Amsterdam, Netherlands · 57+54 other locations
View company
Current open roles at Nebius B.V. on JobCrawls
LocationActive listings
Remote - Global398
Remote - Europe126
Amsterdam, Netherlands57
London, United Kingdom22
Remote - United States20
Remote - Finland18
Berlin, Germany16
Mäntsälä, Finland12
Helsinki, Finland10
Lappeenranta, Finland10
Prague, Czech Republic8
Israel6
Amsterdam5
United Kingdom5
Canada4
Tel Aviv, Israel3
Singapore3
Remote3
Abu Dhabi2
New York City, United States2
London2
Paris, France2
Austin, United States2
Dubai2
France, Paris2
Canada, Remote - United States1
East London, United Kingdom1
Singapore, Singapore1
Minnesota, United States1
Remote - Israel1
Munich, Germany1
UK1
Berlin1
New Jersey, United States1
Remote - Germany1
Abu Dhabi, Dubai1
Prague1
New Jersey, US1
Remote - United Kingdom1
Oklahoma, United States1
Finland1
California, United States1
Abu Dhabi, United Arab Emirates1
Czechia1
Alabama, US1
Béthune, France1
Netherlands1
Austin, Texas1
San Francisco Bay Area, United States1
Kansas City, United States1
Béthune, Pas-de-Calais, France1
Philadelphia, United States1
Dallas, United States1
London, UK1
Israel, Israel1
Remote - EU1
Paris1
Current role mix at Nebius B.V. on JobCrawls
Role typeActive listings
Backend Engineer308
Software Engineer82
Account Executive54
Site Reliability Engineer6
Technical Product Manager5
Technical Project Manager5
Technical Program Manager5
Product Manager4
Sales Representative4
System Engineer4
Data Center Technician3
ML Engineer3
Data Center Operations Technician3
IT Technician2
Product Designer2
Hypervisor Engineer2
Delivery Manager2
Backend engineers, Frontend engineers, Site reliability engineers2
Open Positions at Nebius2
Applied AI Researcher2
Head of Channel Marketing1
Forward Deployment Engineer1
Site Selection & Colocation Manager1
Network Engineer1
DC IT Support Manager1
Application Security Engineer1
MEP Engineer1
Datacenter IT Technician1
Partner Solutions Architect1
Customer Engineer1
Internal Control Business Partner1
Solutions Partner1
Senior System Engineer1
Senior Technical Program Manager1
Solutions Architecture Leader1
Electrical Engineer1
Data Scientist1
Compensation Analyst1
Network Planning Project Manager1
Data Center IT Manager1
Technical Account Manager1
Security Solutions Engineer1
Cloud Solution Architect1
AI and ISV Partner Business Development Manager1
Structured Cabling Design Engineer1
ML Infrastructure Engineer1
Senior Applied AI Solutions Engineer1
VP of Developer Relations & Community1
Senior Support Engineer1
Operations Specialist1
Project Development Manager1
Backend Developer1
Data Center Operations Manager1
Data Center IT Technician1
Mechanical Design Engineer1
Generalist1
HPC Engineer1
Solutions Architect1
Data Center Electrical Lead1
Pricing Director1
Manager, ML Solutions Architecture1
Senior Technical Project Manager1
Accountant1
Product Growth Analytics Lead1
Group Product Manager1
Infrastructure Security Engineer1
Senior Research Scientist1
GTM Recruiting Manager1
Technical Due Diligence Manager1
Technical Support Engineer1
Data Engineer1
Machine Learning Engineer1
Mechanical Data Center Technician1
Sales Engineer1
Senior Software Developer1
AI/ML Specialist Solutions Architect1
Instructional Designer1
Vendor Security & Standards Manager1
ML Solutions Architect1
Physical Security Systems Technician1
Mechanical Engineer1
Vulnerability Operations Center Lead1
IT Risk and Control Manager1
VP of Strategic Sales1
Human Resources Specialist1
Educational Content Author1
Offensive Security Lead1
Data Center Logistics Specialist1
Financial Controller1
Applied ML Engineer1
Field Technical Lead1
Data Center Facilities Manager1
Principal1
Systems HPC Engineer1
Backend Engineers1
Current role-level mix at Nebius B.V. on JobCrawls
Role levelActive listings
Mid-Level387
Senior67
Manager14
Executive3

Help us improve JobCrawls — sign in to sync saved jobs across devices, or send feedback anytime.