Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AI @ 6
AWS @ 4
Azure @ 4
Datadog @ 4
Distributed Systems @ 6
GCP @ 4
GenAI
Generative AI @ 6
Go @ 4
Kibana @ 4
Observability @ 4
Ruby @ 4
Security
Terraform @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
Coinbase is a remote-first, but not remote-only company. Employees are expected to meet quarterly for in-person working sessions called “surges.”
As a Senior Software Engineer, Core Reliability on the Infrastructure Reliability team within Platform, you will help Coinbase scale by improving reliability, security, and deployment safety across the production environment. The team owns systems that secure service configurations and secrets, reduce customer-facing incidents, and strengthen deployment infrastructure supporting thousands of services and hundreds of daily releases.
Responsibilities
- Own the design and delivery of reliability projects and features that improve resiliency across Coinbase’s service environment in partnership with other engineering teams.
- Partner with critical T0/T1 services to understand architecture, improve scalability, and reduce operational toil.
- Build and enhance systems that securely manage service configurations and secrets at scale.
- Improve canary-based release systems and expand deployment capabilities to support thousands of services and hundreds of daily deployments with fewer incidents.
- Drive reliability best practices and strengthen reliability culture across engineering teams at Coinbase.
Requirements
- 5+ years of software engineering experience designing, building, and maintaining production services in service-oriented architectures.
- Experience with Ruby, Go, Terraform, and cloud platforms such as AWS, GCP, or Azure.
- Demonstrated ability to design and operate reliable, high-throughput, low-latency distributed systems at scale.
- Track record of writing well-tested, production-quality code.
- Experience with observability and monitoring tools such as Kibana and Datadog to debug complex production issues, tune system performance, and reduce incident frequency.
- Experience writing and verbally communicating architecture decisions to cross-functional engineering stakeholders.
- Ability to participate in on-call rotations and respond to issues outside normal business hours.
- Ability to use generative AI responsibly while maintaining human oversight and delivering business-ready outputs.
Compensation
The target annual base salary is $191,100 CAD, excluding equity and bonus. Total compensation may also include equity, bonus eligibility, and medical, dental, and vision benefits.
Additional Information
- Candidates may submit a maximum of three applications within a six-month period.
- Coinbase is an Equal Opportunity Employer.
- Reasonable accommodations are available for applicants with disabilities.
- Applicants are subject to Coinbase’s Candidate Privacy Notice. US applicants also agree to arbitration of disputes.