Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AWS @ 3
Azure @ 3
Change Management
Debugging @ 3
Distributed Systems @ 6
GCP @ 3
Go @ 3
Kubernetes @ 3
Linux @ 3
Networking @ 3
Observability @ 3
Oracle @ 3
Payments
Performance Optimization @ 3
Rust @ 3
Software Development @ 5
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Who We Are
About Stripe
Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Stripe's mission is to increase the GDP of the internet.
About the Team
The Core Infrastructure team develops and maintains the infrastructure used by all product teams to create services that underpin Stripe's business operations. The team defines and executes a vision for industry-leading scale and availability while balancing efficiency, latency, and organizational health.
The team advances distributed services and scales existing technologies in collaboration with internal teams and external partners in the open source community. The Infrastructure organization includes teams responsible for operating system components, databases, caching, high availability, disaster recovery, cloud infrastructure, Linux servers, container orchestration, mesh networking, service discovery, change management, and network edge infrastructure.
Responsibilities
- Design, build, and maintain distributed cloud infrastructure and platform services.
- Work on scaling, automation, reliability, and observability of infrastructure services.
- Operate services, debug issues, and support customers.
- Participate in roadmap planning and prioritization.
- Build control plane services for managing primary database and cache infrastructure.
- Build automation for managing cloud components for compute, cache, and networking.
- Provide a strong internal customer experience for Stripe teams building products on the infrastructure.
Requirements
Minimum Requirements
- 5+ years of professional experience in a software development role.
- Experience building, deploying, and managing infrastructure on a major cloud provider.
- A strong engineering background in building platform services and/or distributed systems at scale, with a solid grasp of underlying operating system primitives.
- Experience developing, maintaining, and debugging distributed systems, including diagnosing low-level resource constraints.
- Experience with operational excellence and modern observability practices, including distributed tracing, structured logging, and system-level metrics.
Preferred Qualifications
- Experience with popular cloud technologies such as AWS, Azure, GCP, or Oracle Cloud.
- Experience with Go or other systems languages, such as Rust, C, or C++.
- Experience with Linux OS internals, performance optimization, and kernel-level troubleshooting.
- Experience working with Kubernetes clusters and low-level container mechanics.
- Experience in networking and traffic systems at scale.
- Experience handling critical incidents for production systems.