Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Data Analysis @ 3
GPU
Jira @ 3
Linux @ 3
Project Management @ 3
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Nebius is building a full-stack AI cloud platform supporting developers and enterprises from data and model training through production deployment. The role supports and maintains large-scale GPU clusters in a data center environment, providing advanced IT hardware troubleshooting and assisting with infrastructure implementation projects. The position is based at Nebius's Oklahoma data center.
Responsibilities
- Troubleshoot server hardware issues, including advanced-level problems involving GPU clusters.
- Report hardware issues to problem management and execute workarounds and solutions.
- Proactively handle L1/L2 support tasks through the ticketing system.
- Update processes and documentation for the IT hardware team.
- Collaborate with the Technical Project Manager on current projects.
- Participate in technical project implementations.
- Oversee third-party contractor activities during project implementations.
- Perform basic network tasks, troubleshooting, and support.
Requirements
- Knowledge of data centers and server equipment.
- Advanced knowledge of IT hardware and practical troubleshooting experience.
- Advanced skills with Unix/Linux operating systems and the command line.
- Experience with equipment monitoring, data analysis, and presentation.
- Basic understanding of project management methodologies.
- Basic understanding of project management tools, including JIRA and Gantt charts.
- Proactiveness and a sense of responsibility.
- High proficiency in spoken and written English.
- Driving license.
Bonus Qualifications
- Advanced knowledge of network equipment and troubleshooting.
- Proven project management experience in the IT infrastructure domain.
- Bachelor's degree in Computer Science.
Compensation
Competitive hourly rates range from $30.00 to $45.00 per hour, based on experience.
Benefits
- Career growth and learning opportunities.
- Flexibility and ownership.
- Collaborative and innovative culture.
- Opportunity to work on impactful AI projects.
- International environment and talented teams.
Nebius is an equal opportunity employer committed to fostering an inclusive and diverse workplace. Applicants must be authorized to work in the country in which they apply and must provide proof of employment eligibility as a condition of hire.
More jobs at Nebius
Director, Solutions Architecture - Enterprise
Nebius · United States
USD 195,800-244,700 per year
Product Marketing Manager - Token Factory
Nebius · United States
USD 117,800-222,200 per year
Data Center Technician
Nebius · Kansas City, United States
USD 30-45 per hour
Manager, ML Solutions Architecture - Token Factory
Nebius · United States
USD 228,000-285,000 per year
Data Center GM
Nebius · United States
USD 200,000-250,000 per year
Similar jobs
Technical Program Manager – Chip System Software
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior Software Engineer - Image and Data Processing Libraries
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Embedded System Software Engineer – Platform Execution Lead
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Technical Program Manager, NVIDIA Metropolis
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Technical Program Manager, Deep Learning Frameworks
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Systems Software Engineer - Fleet Debuggability
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Manager, System Software Engineering - Factory
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Math Libraries Engineer - Direct Sparse Solvers
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year