Solutions Architect

at Nebius
📍 Canada
📍 United States
USD 250,000-320,000 per year
SENIOR
✅ Remote

Tech Stack

AI @ 4 AWS @ 4 Ansible @ 7 Azure @ 4 CUDA @ 4 Cloud Computing @ 8 Communication @ 6 Deep Learning @ 4 GPU @ 4 HPC @ 4 Hiring @ 4 IaC KubeFlow @ 4 Kubernetes @ 4 MLOps Machine Learning @ 4 Marketing OpenCL @ 4 PyTorch @ 4 Python @ 6 Slurm @ 4 TensorFlow @ 4 Terraform @ 7

Details

Nebius is building a full-stack AI cloud platform for developers and enterprises, supporting workloads from data and model training through production deployment. The Customer Experience Team helps customers solve real-world AI and ML challenges at massive GPU cloud scale, working with GPUs such as H200, B200, and GB200 and modern machine learning frameworks.

The role focuses on cloud infrastructure and MLOps. You will design and implement solutions for clients, leverage cloud technologies for ML and AI teams, and serve as a trusted technical advisor for building their pipelines. The position can be performed remotely from the United States or Canada.

Responsibilities

  • Act as a trusted advisor to clients, providing technical expertise and guidance throughout engagements.
  • Conduct proofs of concept, workshops, presentations, and training sessions on GPU cloud technologies and best practices.
  • Understand customer business requirements and develop solution architectures aligned with their needs.
  • Design and document Infrastructure as Code solutions, documentation, and technical how-to materials in collaboration with support engineers and technical writers.
  • Help customers optimize pipeline performance and scalability for efficient use of cloud resources and Nebius AI services.
  • Serve as the primary subject-matter expert for customer scenarios across product, technical support, and marketing teams.
  • Support marketing activities at hackathons, conferences, workshops, webinars, and other events.

Requirements

  • 5–10+ years of experience as a cloud solutions architect, systems or network engineer, developer, or in a similar technical role focused on cloud computing.
  • Strong hands-on experience with Infrastructure as Code and configuration management tools, preferably Terraform and Ansible.
  • Experience with Kubernetes.
  • Ability to write code in Python.
  • Solid understanding of GPU computing practices for ML training and inference workloads.
  • Knowledge of GPU software stack components, including drivers and libraries such as CUDA and OpenCL.
  • Excellent communication skills.
  • Customer-centric mindset.

Additional Qualifications

  • Hands-on experience with HPC or ML orchestration frameworks such as Slurm or Kubeflow.
  • Hands-on experience with deep learning frameworks such as TensorFlow or PyTorch.
  • Understanding of the cloud ML tools landscape from industry leaders including NVIDIA, AWS, Azure, and Google.

Benefits

  • 100% company-paid medical, dental, and vision coverage for employees and families.
  • 401(k) plan with up to a 4% company match and immediate vesting.
  • 20 weeks of paid parental leave for primary caregivers and 12 weeks for secondary caregivers.
  • Remote work reimbursement of up to $85 per month for mobile and internet expenses.
  • Company-paid short-term disability, long-term disability, and life insurance.
  • Career growth and learning opportunities.
  • Flexibility and ownership.
  • Collaborative and innovative culture.
  • Opportunity to work on impactful AI projects.

Compensation

The on-target earnings range is $250,000–$320,000 USD. Actual compensation depends on job-related factors including experience, skills, qualifications, hiring level, and geographic location.

Nebius is an equal opportunity employer committed to an inclusive and diverse workplace. Applicants must be authorized to work in the country in which they apply and must provide proof of employment eligibility as a condition of hire.

More jobs at Nebius

Similar jobs