Principal Machine Learning Engineer, Accelerated Apache Spark
at Nvidia
USD 272,000-431,200 per year
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Algorithms
CUDA @ 3
Data Science @ 4
ETL
GPU
GenAI @ 7
Java @ 4
LLM @ 7
Machine Learning @ 4
Pandas @ 6
PyTorch @ 6
Python @ 6
Reinforcement Learning @ 7
SQL
Scala @ 4
Spark @ 4
TensorFlow @ 6
XGBoost @ 7
scikit-learn @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team.
Apache Spark is the most popular data processing engine in data centers for running large scale workloads for ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and use cases. NVIDIA GPUs offer a promising avenue for significantly speeding up and/or lowering the cost of running Apache Spark applications at massive scales. You will work with the open source community to accelerate Apache Spark with GPUs. You will apply the latest ML/AI methods to empower enterprises to migrate Spark workloads onto GPUs at scale.
Responsibilities
- Design and implement machine learning solutions for performance prediction and optimization of GPU accelerated enterprise Apache Spark workloads.
- Develop advanced algorithms and adaptive systems to continuously improve the performance of Apache Spark workloads on GPUs.
- Develop AI-based agents and tools to assist with fixing system issues and application optimization.
- Collaborate with key partners and customers on the deployment of complex machine learning solutions in various environments.
- Maintain deep domain expertise by knowing the latest published advances in ML systems and algorithms.
- Provide technical mentorship and leadership in data science and machine learning to a team of engineers.
Requirements
- BS, MS, or PhD or equivalent experience in Machine Learning, Data Science, Computer Science or a closely related field.
- 12+ years of professional experience in designing, implementing, and productionizing high-quality ML/DL solutions.
- 5+ experience as technical lead in ML model development.
- Proven hands-on experience (2+ years) with large-scale data processing platforms, such as Apache Spark.
- Proven ability to employ modern tooling and sound techniques for all aspects of crafting, deploying, and maintaining machine learning models.
- Excellent programming skills in Python and Python data science related libraries like numpy, pandas, scikit-learn, scipy, pytorch, and tensorflow.
- Deep experience with sophisticated ML methodologies, including LLM/GenAI, reinforcement learning, and adaptive, on-line ML systems.
- Strong expertise in feature engineering, feature importance assessment, and developing boosted tree model solutions (e.g., XGBoost).
Ways to stand out from the crowd
- Understanding of the internal workings and architecture related to Apache Spark.
- Familiarity with NVIDIA GPUs and CUDA.
- Experience coding in Scala, Java, and/or C++.
More jobs at Nvidia
Senior System Software Engineer - Halos Core And Robotics Platform
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Director, Autonomous Vehicles Platform
Nvidia · Santa Clara, United States
USD 320,000-488,800 per year
Senior Developer Technology Engineer - Edge Agentic Ai
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior System Software Engineer - Halos
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Engineering Manager, Drive Os Communication Infrastructure
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Similar jobs
Principal Data Scientist - Cloud Gaming And Ai
Nvidia · Santa Clara, United States
USD 248,000-379,500 per year
Senior Machine Learning Engineer, Relevance And Personalization (Query Intelligence)
Airbnb · United States
USD 200,000-235,000 per year
Principal AI/ML Researcher and Engineer
Airbnb · United States
USD 296,000-370,000 per year
Senior AI Security Researcher
Nvidia · Durham, United States
USD 224,000-431,200 per year
Senior Software Engineer - Python Numerical Computing Libraries
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Senior Software Engineer, Ai Networking
Nvidia · Santa Clara, United States
USD 152,000-287,500 per year
Senior Staff Machine Learning Engineer, Trust
Airbnb · United States
USD 244,000-305,000 per year
AI/Ml Specialist Solutions Architect
Nebius · United States, Canada
USD 250,000-320,000 per year