Senior MLOps Engineer - DSX Enablement

at Nvidia
📍 Germany
PLN 292,500-650,000 per year
SENIOR
✅ Remote

Tech Stack

AI @ 4 Bash @ 6 CI/CD @ 3 CUDA @ 4 Communication @ 6 Data Engineering @ 7 Data Science @ 7 Deep Learning @ 4 Go @ 6 InfiniBand @ 4 Kubernetes @ 4 LLM Leadership @ 6 Linux @ 4 MLOps @ 3 Machine Learning @ 4 NVLink @ 4 Networking @ 4 Observability @ 3 Python @ 6 Rust @ 6 Security @ 4

Details

NVIDIA is seeking a Senior MLOps Engineer to join the DSX Enablement team, collaborating closely with strategic customers to implement and enhance AI workloads. The team partners with innovative AI companies and open-source communities to address challenging technical problems.

Responsibilities

  • Build and deploy custom AI solutions on NeoCloud platforms and NVIDIA Cloud Partners, including distributed training, inference optimization, and MLOps pipelines.
  • Act as a primary technical contact for internal and external customers and partners, guiding joint engagements, ensuring the success of initiatives on DGX Cloud, and solving complex production problems.
  • Collaborate with teams building infrastructure software and accelerated frameworks for AI applications.
  • Profile and tune large-scale training and inference workloads on NVIDIA Cloud Partner platforms, reducing latency, cost, and operational risk.
  • Develop open-source tools and reference architectures for building and managing machine learning and AI workloads, pipelines, and systems at scale.
  • Support LLM performance evaluation and new hardware in open-source frameworks.

Requirements

  • BS, MS, or Ph.D. in Computer Science, Computer/Electrical Engineering, or a related technical field, or equivalent experience.
  • 8+ years of experience in technical roles such as data science, data engineering, or ML engineering, ideally focused on large-scale production systems.
  • Demonstrated AI/ML experience across multiple phases of the machine learning lifecycle, from exploratory analysis through production systems.
  • Experience with Linux, batch schedulers, Kubernetes, distributed filesystems, and advanced networking at datacenter scale.
  • Scripting and programming skills in Bash and Python, along with systems programming skills in C++, Go, or Rust.
  • Experience using machine learning or deep learning frameworks for training and inference.
  • Excellent communication and technical presentation skills, including the ability to articulate architectures, trade-offs, and recommendations to engineering and leadership audiences.
  • A record of engineering discipline and execution on individual and collaborative projects.

Preferred Qualifications

  • Experience contributing to and working in open-source communities.
  • Experience with the NVIDIA ecosystem, including DGX systems, CUDA, NeMo, RAPIDS, Triton, NIM, InfiniBand, NVLink, and RoCE.
  • Experience building machine learning systems in security-critical environments.
  • Experience with distributed training and inference frameworks.
  • Familiarity with cloud-native MLOps practices, including containerization, CI/CD pipelines, workflow automation, observability stacks, and GitOps workflows.
  • Deep systems knowledge for diagnosing and fixing performance or correctness issues across hardware, networking, accelerators, hypervisors or operating systems, compilers or runtimes, application code, and libraries.

Benefits

NVIDIA offers competitive salaries and a generous benefits package.

Compensation

For Poland, the base salary range is 292,500 PLN–507,000 PLN for Level 4 and 375,000 PLN–650,000 PLN for Level 5. Base salary is determined by location, experience, and the pay of employees in similar positions.

More jobs at Nvidia

Similar jobs