Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Airflow @ 4
Audit @ 4
Change Management
Distributed Systems @ 4
Engineering Management @ 6
Go @ 4
Hadoop @ 4
Java @ 4
Microservices
Networking @ 4
Observability @ 4
Python @ 4
Ruby @ 4
Security @ 4
Software Development @ 8
Spark @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Stripe's Infrastructure organization builds and operates shared platforms that enable engineering teams to build, deploy, run, and scale reliable products. The organization owns foundational capabilities across service networking and discovery, safe feature and configuration changes, data storage and orchestration, and batch computing.
The platforms support a global engineering audience and operate at significant scale, including secure service-to-service communication across tens of thousands of microservices, service and feature deployment infrastructure, containerized service infrastructure, and data platforms for storage, pipeline orchestration, and large-scale processing. Batch compute platforms support Apache Airflow, Apache Spark, Apache Iceberg, Apache Hadoop, and Apache Celeborn.
As an Engineering Manager, you will lead or strongly influence teams working on business-critical infrastructure. You will balance reliability, scalability, security, developer experience, and operational rigor while partnering with product, security, data, support, and adjacent infrastructure leaders.
Responsibilities
- Drive impact across multiple teams or a significant infrastructure area independently, translating organizational goals into execution plans and delivering results across parallel workstreams.
- Define strategy and roadmaps while anticipating internal user needs, technical challenges, reliability risks, and business priorities.
- Coach senior engineers and emerging leaders, create stretch opportunities, and provide candid feedback.
- Align product, security, data, support, and engineering stakeholders, resolving ambiguity and cross-team dependencies.
- Recruit, develop, and retain senior engineering talent while building an inclusive, high-performing organization.
- Set high standards for reliability, availability, scalability, incident response, and engineering quality.
- Partner with Staff engineers on architecture decisions involving distributed systems, service discovery, configuration propagation, production change management, data processing, and internal platform design.
Requirements
Minimum Requirements
- Bachelor's degree or equivalent practical experience and 10+ years of software development experience.
- 4+ years of engineering management experience with demonstrated autonomous impact.
- Experience leading high-complexity engineering teams delivering multiple difficult, high-impact workstreams in parallel.
- Ability to influence and align cross-functional stakeholders without direct authority.
- Meaningful recruiting experience, including attracting and developing senior engineering talent.
- Technical depth sufficient to guide architecture decisions and maintain high standards for product quality, reliability, and engineering craft.
Preferred Qualifications
- Experience building or operating high-availability distributed systems, internal platforms, or developer infrastructure used by broad engineering audiences.
- Experience with service networking, service discovery, service mesh, feature flags, configuration management, progressive delivery, deployment tooling, or production change governance.
- Experience with data infrastructure, including large-scale storage, data orchestration, batch compute, or Apache Airflow, Apache Spark, Apache Iceberg, Apache Hadoop, or Apache Celeborn.
- Experience improving reliability through observability, health checks, incident response, automated recovery, capacity planning, or reducing the impact of production changes.
- Experience designing secure access controls, audit systems, approval workflows, or governance mechanisms for production infrastructure.
- Experience with cloud infrastructure and distributed systems technologies such as Amazon Web Services, DynamoDB, Aurora, S3, Ruby, Go, Java, or Python.
- A product mindset for internal platforms, including empathy for internal users and a focus on self-service and safe workflows.
Office Expectations
Office-assigned Stripes in most locations are expected to spend at least 50% of their time each month in their local office or with users. Expectations may vary by role, team, and location and will be discussed during the hiring process.