Principal Site Reliability Engineer, Platform Engineering: Dedicated
Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Change Management
Communication @ 6
Compliance
Distributed Systems @ 7
Go @ 7
IaC
Leadership @ 6
Observability @ 4
Python @ 7
Ruby @ 7
Security
Technical Leadership @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
GitLab is seeking a Principal Engineer with deep expertise in Site Reliability, Backend, or Platform Engineering to help shape the next phase of GitLab Dedicated, its fully managed single-tenant SaaS offering. The role will set direction for scaling isolated, customer-specific environments while maintaining reliability, security, and compliance.
The position will lead platform and operating-model transformation across resilience and failover, tenant orchestration, change management, self-service tooling, and platform integrations. It will also help align GitLab Dedicated with modular and cell-based architectures, strengthen service ownership, and establish scalable patterns that reduce operational complexity.
Responsibilities
- Set technical direction for GitLab Dedicated, shaping architecture and platform strategy as the fleet of isolated, single-tenant environments grows.
- Lead platform transformations across resilience, failover, tenant orchestration, change management, self-service tooling, and platform integrations.
- Drive scalable, modular architecture aligned with GitLab’s broader Cells strategy while preserving security, isolation, and compliance requirements.
- Strengthen service ownership and operational maturity, helping engineering teams build, operate, and improve the production systems they own.
- Identify and address systemic reliability and scalability risks using production signals, incident patterns, and architectural insight.
- Establish reusable platform patterns and automation that reduce operational toil and enable efficient scaling.
- Lead complex technical decisions across teams, balancing reliability, security, cost, maintainability, and customer needs.
- Advance engineering excellence through architectural leadership, mentorship, and influence with senior engineers and engineering leaders.
Requirements
- Deep expertise in Site Reliability, Platform, Infrastructure, or Backend Engineering, with experience designing and operating large-scale production systems.
- Hands-on experience with cloud infrastructure, automation, observability, infrastructure as code, and modern production engineering practices.
- Strong software engineering fundamentals, with experience building production systems or infrastructure tooling in Go, Ruby, Python, or similar languages.
- Strong distributed systems and systems-design expertise, including reliability, failure isolation, scalability, and operational complexity.
- Track record of technical leadership across multiple teams, setting direction and driving complex initiatives through influence.
- Experience leading significant platform or infrastructure transformations, including modernization, modularization, or scaling systems through major growth.
- Experience improving how engineering teams own and operate production systems, strengthening reliability, operational readiness, and accountability at scale.
- Exceptional technical communication and influence, with the ability to build alignment, mentor senior engineers, and guide complex architectural decisions.
Compensation
- United States salary range: $223,200–$380,400 USD per year.
- The base salary range excludes bonuses, equity, and benefits.
Benefits
- Benefits supporting health, finances, and well-being.
- Flexible paid time off.
- Team Member Resource Groups.
- Equity compensation and Employee Stock Purchase Plan.
- Growth and Development Fund.
- Parental leave.
GitLab is a remote organization. Location-based eligibility requirements may apply, and the Talent Acquisition team can provide additional information during the recruiting process. GitLab is an equal opportunity workplace and affirmative action employer.