Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 3
Agile @ 2
GPU @ 3
Networking @ 3
Project Management @ 3
Team Management @ 3
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Nebius is building a full-stack AI cloud platform supporting developers and enterprises from data and model training through production deployment. The role supports modern GPU clusters and data center infrastructure, including GPU servers, high-speed networks, compute systems, and storage systems. You will act as a link between stakeholders, project teams, and technical teams while driving operational activities.
Responsibilities
- Manage and deliver IT infrastructure projects and change requests in data centers.
- Proactively identify and resolve technical and operational issues.
- Manage risks and dependencies in live environments.
- Maintain operational documentation, including project plans, risk logs, issue trackers, and deployment reports.
- Lead stakeholder communication, provide progress updates, escalate blockers, and manage expectations.
- Lead meetings and technical discussions focused on decision-making, clear next steps, and timely execution.
- Collaborate with cross-functional teams.
- Work directly with data center engineers, vendors, and operations teams.
- Participate in automation, process improvement, and design support tasks as needed.
Requirements
- Proven experience in data center infrastructure operations, preferably in high-performance or AI-oriented environments.
- Solid understanding of data center design and operations, including racking, cabling, power planning, and cooling considerations.
- Hands-on experience working with IT hardware in data centers.
- Comfort working in data center environments and during live operations.
- Experience in project delivery, team management, and cross-team collaboration.
- Strong problem-solving skills and the ability to troubleshoot technical and workflow issues in an IT stack.
- Ability to document processes, maintain tracking systems, and produce clear deployment reports.
- Proactive, responsible, and purposeful approach to work.
- Willingness to take occasional business trips.
Preferred Qualifications
- Familiarity with structured project management or delivery frameworks such as PRINCE2, PMI, or Agile.
- Relevant technical certifications.
- In-depth technical knowledge of GPU servers, compute nodes, and high-speed networking.
Benefits
- Base compensation range of $147,200–$183,900 USD.
- Competitive compensation and benefits.
- Career growth and learning opportunities.
- Flexibility and ownership.
- Collaborative and innovative culture.
- Opportunity to work on impactful AI projects.
- International environment and talented teams.
Nebius is an equal opportunity employer committed to an inclusive and diverse workplace. Applicants must be authorized to work in the country in which they apply and must provide proof of employment eligibility as a condition of hire.
More jobs at Nebius
Director, Solutions Architecture - Enterprise
Nebius · United States
USD 195,800-244,700 per year
Product Marketing Manager - Token Factory
Nebius · United States
USD 117,800-222,200 per year
Data Center Technician
Nebius · Kansas City, United States
USD 30-45 per hour
Manager, ML Solutions Architecture - Token Factory
Nebius · United States
USD 228,000-285,000 per year
Data Center GM
Nebius · United States
USD 200,000-250,000 per year
Similar jobs
Technical Project Manager / IT Infrastructure Engineer
Nebius · United States
USD 147,200-183,900 per year
Technical Program Manager – Chip System Software
Nvidia · Santa Clara, United States
USD 200,000-322,000 per year
Senior Technical Project Manager – Applied AI
Nebius · Palo Alto, United States
USD 147,200-224,000 per year
Senior System Software Engineer, Agentic Inference – Dynamo
Nvidia · Santa Clara, United States
USD 224,000-431,200 per year
Senior Technical Program Manager, NVIDIA Metropolis
Nvidia · Santa Clara, United States
USD 168,000-322,000 per year
Senior Math Libraries Engineer - Direct Sparse Solvers
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Principal Architect, System Software - Orbital Data Center
Nvidia · Santa Clara, United States
USD 272,000-431,200 per year
Manager, Software Architecture
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year