Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
API
AWS @ 4
Audit @ 4
Azure @ 4
Change Management
Distributed Systems @ 6
IaC
Kafka @ 3
Kubernetes @ 4
Observability @ 3
Security
Terraform @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Stripe's Core Change Management group owns the systems that allow engineers to ship code, configuration, and infrastructure changes safely and quickly. The role is embedded primarily on the Service Deployments team, which owns Stripe's end-to-end code deployment platform and systems in the critical path of every engineer's workflow.
Responsibilities
- Own end-to-end technical delivery of large, ambiguous infrastructure projects from design through production launch and long-term reliability.
- Architect the next generation of Stripe's deployment platform, including multi-service dependency-aware autodeploy pipelines, Kubernetes-native deployment primitives, and fleetwide container migration.
- Define API contracts, rollout strategies, and operational models used by hundreds of teams.
- Extend deploy anomaly detection, including earlier traffic-split stages, API- and method-based regression detection, and self-service onboarding.
- Lead the migration of host-based services to containerized, fleetwide deployments while maintaining production operations.
- Own reliability and operational excellence, lead incident response, reduce operational toil, and prioritize reliability, security, and maintainability.
- Build deployment event infrastructure, including event schemas, durability models, notification systems, and integration contracts for observability and automation.
- Collaborate with Resource Automation on IAM, account provisioning, infrastructure automation, and cloud resource management.
- Collaborate with Feature Deployments on feature flags, configuration management, change audit logs, and change-safety tooling.
- Work with the service mesh team on canary rollouts, weighted routing, and merchant-priority traffic shaping.
- Lead design reviews, establish deployment safety and developer experience standards, mentor senior engineers, and provide technical direction on complex platform challenges.
Requirements
- 10+ years of professional software engineering experience designing and shipping large-scale production infrastructure systems.
- Experience leading large, ambiguous infrastructure projects end-to-end, including cross-team dependencies and migrations across many consuming teams.
- Deep expertise in distributed systems and deployment orchestration, including service scheduling, staged delivery, rollout strategies, and failure modes.
- Hands-on experience with Kubernetes and container-based deployments, service lifecycle management, workload scheduling, and VM-to-container fleet migrations.
- Strong background in service reliability and operational excellence, including incident response, toil reduction, and building reliable, debuggable, and maintainable systems.
- Broad technical impact across multiple large systems, including complex codebases, code review, mentorship, and setting technical direction.
Preferred Requirements
- Experience with deployment safety systems such as anomaly detection, automated rollback, or progressive delivery.
- Familiarity with event-driven architectures such as Kafka for deployment lifecycle observability and notification.
- Experience with Infrastructure as Code at scale, including Terraform or equivalent, and cloud resource governance and IAM in AWS or Azure.
- Experience building developer platforms or internal tooling with a strong focus on developer experience.
- Experience with feature flags, configuration distribution, or audit-log infrastructure.
- Familiarity with service mesh concepts, canary deployments, weighted routing, and traffic splitting.
In-Office Expectations
Office-assigned Stripes in most locations are expected to spend at least 50% of each month in their local office or with users. This expectation may vary by role, team, and location.
More jobs at Stripe
Software Engineer, Dev Productivity
Stripe · Taipei, Taiwan
TWD 2,301,000-3,451,600 per year
Backend Engineer, Developer & End-User Experience Platform
Stripe · South San Francisco, United States, New York City, United States
USD 190,400-285,600 per year
Staff Software Engineer, Payments Intelligence
Stripe · Seattle, United States
USD 224,000-336,000 per year
Staff Software Engineer, Financial Crimes
Stripe · Toronto, Canada
CAD 208,000-312,000 per year
Full Stack Engineer, Expansion
Stripe · Toronto, Canada
CAD 135,200-202,800 per year
Similar jobs
Senior Software Engineer
SentinelOne · United States
USD 132,000-182,000 per year
Staff Software Engineer, Customer Administration
Coinbase · India
INR 9,424,500 per year
Forward Deployed Engineer - Physical AI Cloud Platform
Nebius · United States, Austin, United States
USD 179,500-224,300 per year
Software Engineer, Product Security Data Platforms
Stripe · Seattle, United States
USD 156,800-235,200 per year
Senior Manager, Platform Operations
Collibra · United States
USD 168,000-210,000 per year
Staff + Senior Software Engineer, Cloud Inference Launch Engineering
Anthropic · San Francisco, United States
USD 320,000-485,000 per year
Senior Software Engineer - Public Cloud Engineering
Bloomberg · New York City, United States
USD 160,000-240,000 per year
Staff + Senior Software Engineer, Cloud Inference
Anthropic · San Francisco, United States
USD 320,000-485,000 per year