Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI
Data Analysis @ 3
GPU
Jira @ 3
Linux @ 3
Project Management @ 3
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Nebius is building a full-stack AI cloud platform supporting developers and enterprises from data and model training through production deployment. The role supports and maintains large-scale GPU clusters in a data center environment, providing advanced IT hardware troubleshooting and assisting with infrastructure implementation projects. The position is based at Nebius's Oklahoma data center.
Responsibilities
- Troubleshoot server hardware issues, including advanced-level problems involving GPU clusters.
- Report hardware issues to problem management and execute workarounds and solutions.
- Proactively handle L1/L2 support tasks through the ticketing system.
- Update processes and documentation for the IT hardware team.
- Collaborate with the Technical Project Manager on current projects.
- Participate in technical project implementations.
- Oversee third-party contractor activities during project implementations.
- Perform basic network tasks, troubleshooting, and support.
Requirements
- Knowledge of data centers and server equipment.
- Advanced knowledge of IT hardware and practical troubleshooting experience.
- Advanced skills with Unix/Linux operating systems and the command line.
- Experience with equipment monitoring, data analysis, and presentation.
- Basic understanding of project management methodologies.
- Basic understanding of project management tools, including JIRA and Gantt charts.
- Proactiveness and a sense of responsibility.
- High proficiency in spoken and written English.
- Driving license.
Bonus Qualifications
- Advanced knowledge of network equipment and troubleshooting.
- Proven project management experience in the IT infrastructure domain.
- Bachelor's degree in Computer Science.
Compensation
Competitive hourly rates range from $30.00 to $45.00 per hour, based on experience.
Benefits
- Career growth and learning opportunities.
- Flexibility and ownership.
- Collaborative and innovative culture.
- Opportunity to work on impactful AI projects.
- International environment and talented teams.
Nebius is an equal opportunity employer committed to fostering an inclusive and diverse workplace. Applicants must be authorized to work in the country in which they apply and must provide proof of employment eligibility as a condition of hire.
More jobs at Nebius
Senior Product Manager, Enterprise
Nebius · United States
USD 179,500-224,300 per year
Senior Enterprise Applications Engineer
Nebius · United States
USD 147,200-183,900 per year
Senior Technical Program Manager
Nebius · United States
USD 115,000-275,000 per year
Director of Product, Ecosystem
Nebius · United States
USD 228,000-285,000 per year
Technical Program Manager (Post-Sales)
Nebius · United States
USD 179,500-224,300 per year
Similar jobs
Team Lead, Human Data Operations - Post-Training
SpaceXAI · United Kingdom, Indonesia, Ireland, India, Japan, South Korea, Philippines, United States, Dubai, United Arab Emirates, Singapore, Singapore
USD 104,000-156,000 per year
Expert Team Lead, Human Data Operations
SpaceXAI · United Kingdom, Indonesia, Ireland, India, Japan, South Korea, Philippines, United States, Dubai, United Arab Emirates, Singapore, Singapore
USD 104,000-170,400 per year
Software Program Manager, New Product Introduction
Nvidia · Santa Clara, United States
USD 108,000-212,800 per year
Compiler Verification Engineer, Compute Performance – GPU
Nvidia · Austin, United States
USD 140,000-224,200 per year
Technical Program Manager – Chip System Software
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior Software Engineer - Image and Data Processing Libraries
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Embedded System Software Engineer – Platform Execution Lead
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Technical Program Manager, NVIDIA Metropolis
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year