Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AWS @ 7
ArgoCD @ 4
Azure @ 7
CI/CD @ 4
Distributed Systems @ 4
GCP @ 7
GitHub @ 4
GitHub Actions @ 4
Go @ 7
IaC
Jenkins
Kafka @ 7
Kubernetes @ 7
Observability
Python @ 7
Redis @ 7
Terraform @ 4
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
SentinelOne is seeking a Staff Infrastructure Engineer to build and operate infrastructure supporting a unified autonomous cybersecurity platform. The role focuses on data platforms that operate with 99.99% uptime, ingest petabytes of data daily, process more than 2 trillion events, and return queries with p95 latencies below 5 seconds.
Due to Federal Government contract requirements, U.S. Citizenship is required. FedRAMP staff may be subject to customer or third-party background checks up to and including Secret Clearance if required by their role.
Responsibilities
- Lead the design and operation of distributed data services, including self-hosted Kafka and Redis, at massive scale across Kubernetes clusters and multi-cloud environments.
- Build highly automated, self-service infrastructure that runs across AWS, GCP, and air-gapped on-premises environments.
- Manage data infrastructure supporting 5 or more PB per day of ingestion, ensuring low-latency, high-throughput, and cost-effective operation at global scale.
- Consolidate and optimize multi-tenant Kafka clusters to reduce cost, improve resilience, and streamline operations.
- Drive Redis and Kafka lifecycle automation using GitOps principles, including ArgoCD and Terraform.
- Define and implement standards for observability, high availability, backup, and disaster recovery of stateful Kubernetes workloads.
- Partner with FinOps and engineering stakeholders to optimize performance, cost, and operational overhead across data platform components.
- Own the end-to-end platform experience for mission-critical open-source systems serving hundreds of product teams.
- Collaborate with teams in Europe and India. The U.S. Eastern Time Zone is preferred due to collaboration requirements.
- Work with Kubernetes, including EKS and GKE, Jenkins, GitHub Actions, ArgoCD, and Terraform.
Requirements
- 8 or more years of experience in infrastructure or platform engineering, with a proven track record operating stateful distributed systems at scale.
- Deep hands-on experience with self-hosted Kafka and Redis running in Kubernetes, including performance tuning, scaling, partitioning, persistence, and operator-based lifecycle management.
- Strong understanding of Kubernetes internals and best practices for managing stateless and stateful production workloads.
- Experience providing Database-as-a-Service or Messaging-as-a-Service for internal development teams or external customers.
- Exposure to multi-cloud environments and strong expertise in at least one major cloud provider: AWS, GCP, or Azure.
- Experience with Infrastructure as Code and GitOps practices using Terraform, ArgoCD, or Pulumi.
- Familiarity with blue-green, canary, and rolling deployment strategies.
- Strong scripting or development skills using Python, Go, or similar technologies.
- Solid understanding of CI/CD pipelines and workflow automation, including GitHub Actions and Argo Workflows.
- U.S. Citizenship is required.
Benefits
- Restricted Stock Units and Employee Stock Purchase Plan.
- Flexible time off, paid company holidays and sick time, gender-neutral parental leave, and grandparent leave.
- Medical, dental, and vision coverage; 401(k) with company match; life and disability insurance; FSAs; voluntary benefits; employee assistance; prepaid legal services; pet insurance; cancer care; and global business travel medical insurance.
- Home office allowance and mobile phone reimbursement.
- Wellness coach, wellness or gym reimbursement, fertility coverage, and adoption and surrogacy reimbursement.
Compensation
The U.S. base salary range is $156,000–$215,000 per year. The range may vary based on the candidate's location, and a different range may apply in some locations.
More jobs at SentinelOne
Senior Software Engineer, Agent Platform
SentinelOne · United States
USD 132,000-182,000 per year
Director of Product Management
SentinelOne · United States
USD 206,200-309,000 per year
Solutions Engineer
SentinelOne · United States
USD 196,000-250,000 per year
Senior Manager of Solution Engineering, Enterprise
SentinelOne · United States
USD 232,000-319,000 per year
Staff Endpoint Software Engineer
SentinelOne · United States
USD 156,000-215,000 per year
Similar jobs
Senior DevOps Engineer, Platform Engineering
Nvidia · Santa Clara, United States
USD 176,000-276,000 per year
Staff Infrastructure Engineer
SentinelOne · United States
USD 132,000-215,000 per year
Senior Software Engineer, Analyst Experience Features
SentinelOne · United States
USD 151,800-209,300 per year
Infrastructure Security Engineer
SpaceXAI · Palo Alto, United States, Washington, United States, Austin, United States, New York City, United States
USD 100,000-258,000 per year
Staff Backend Software Engineer, Agent Platform
SentinelOne · United States
USD 156,000-215,000 per year
Senior Systems Software Engineer, Developer Productivity and Cloud Automation - GeForce NOW
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
Operations Engineer, BizTech
Airbnb · United States
USD 136,000-160,000 per year
Senior Software Engineer, AI Inference Systems
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year