Machine Learning Engineer, AI Safety

at Nvidia
USD 124,000-241,500 per year
MIDDLE
✅ On-site

Tech Stack

AI @ 3 Algorithms @ 6 Communication @ 3 Compliance @ 3 GenAI Generative AI @ 3 LLM MLOps Machine Learning @ 3 PyTorch @ 3 Python @ 3 RAG Security @ 3

Details

NVIDIA is developing AI-based products across multiple domains and works with AI companies as partners and customers. The team focuses on content safety, ML fairness, robustness, explainability, and security for generative language and multimodal models. This role will assess, quantify, and improve the safety and inclusivity of large language models at scale across research and production engineering teams.

Responsibilities

  • Develop datasets and models for training and evaluating models and end-to-end systems for content safety, product security, robustness, and ML fairness.
  • Research and implement techniques for bias detection and mitigation in LLMs and systems using LLMs, including retrieval-augmented generation (RAG) systems.
  • Define and track key metrics for responsible LLM behavior and usage.
  • Follow MLOps best practices for automation, monitoring, scalability, and safety.
  • Contribute to the MLOps platform and develop safety tools to improve the effectiveness of ML teams.
  • Collaborate with engineers, data scientists, and researchers to develop solutions to content safety and ML fairness challenges.

Requirements

  • Master’s or PhD in Computer Science, Electrical Engineering, or a related field, or equivalent experience.
  • At least 2 years of experience developing and deploying machine learning models in production.
  • Strong understanding of machine learning principles and algorithms.
  • Hands-on programming experience with Python and in-depth knowledge of machine learning frameworks such as Keras or PyTorch.
  • At least 1 year of experience in one or more of content safety, ML fairness, robustness, AI model security, or related areas.
  • Experience in content safety areas such as hate and harassment, sexualized content, harmful or violent content, or other related areas.
  • Experience working with large multimodal datasets and multimodal models.
  • Strong problem-solving and analytical abilities.
  • Excellent collaboration and communication skills.
  • Demonstrated behaviors that build trust, including humility, transparency, respect, and intellectual honesty.

Preferred Qualifications

  • Experience aligning or fine-tuning LLMs, including regular LLMs, vision-language models (VLMs), or any-to-text models.
  • Experience with multimodal or multilingual content safety, legal requirements, and regulatory compliance.
  • Knowledge of robustness issues including hallucinations, digressions, and generative misinformation.
  • Experience with generative AI security, including prompt stability, model extraction, confidentiality and data extraction, integrity, availability, and adversarial robustness.
  • Passion for AI and a demonstrated commitment to advancing the field through innovative research, scientific research, or publications.

Compensation and Benefits

The base salary range is USD 124,000–195,500 for Level 2 and USD 152,000–241,500 for Level 3. Salary is determined by location, experience, and pay for similar positions. The role also includes eligibility for equity and benefits.

Applications will be accepted at least until July 28, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes and is committed to an inclusive work environment and equal opportunity employment.

More jobs at Nvidia

Similar jobs