Tech Stack
Tag name is followed by "@" symbol and proficiency level value.
About proficiency levels:
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
AWS @ 6
Azure @ 6
CI/CD @ 4
Communication @ 7
Distributed Systems @ 7
GCP @ 6
Grafana @ 4
Kubernetes @ 4
Linux @ 4
Observability @ 4
Performance Optimization
Prometheus @ 4
RAG
Terraform @ 6
- 1-2 — basic awareness. Minimal hands-on experience, and a rudimentary understanding of the technology's purpose;
- 3-6 — daily use. Comfortable and regular usage, capable of handling common tasks and challenges related to the technology;
- 7-9 — you are an expert, you can teach others, you know all the pitfalls and tricks;
- 10 — exceptional knowledge, comprehensive understanding, and adeptness in all aspects of the technology, including advanced problem-solving. Think twice before claiming or demanding such level.
Details
The Database Systems team specializes in high-performance distributed databases. The team built Rockset, the real-time search, analytics, and vector database that powers vector search and retrieval-augmented generation (RAG) at OpenAI. Rockset also supports core functionality across OpenAI's product lines and critical internal use cases.
The role focuses on distributed systems, close-to-the-metal performance optimization, and building scalable database infrastructure. The core engine is written in C++. Engineers contribute to ingestion, query execution, indexing, and storage while improving online database reliability and throughput.
Responsibilities
- Design, build, and operate high-performance distributed systems.
- Identify and resolve performance bottlenecks to scale infrastructure to the next order of magnitude.
- Define long-term technical direction and guide system evolution.
- Collaborate with product, engineering, and research teams to deliver scalable and reliable infrastructure.
- Investigate complex production issues across the stack.
- Contribute to incident response, postmortems, and system reliability best practices.
Requirements
- 4+ years of relevant industry experience, including 2+ years leading large-scale, complex projects or teams as an engineer or tech lead.
- Experience building, scaling, and optimizing distributed systems with a strong focus on performance, reliability, and scalability.
- Strong communication skills and the ability to collaborate across highly technical and cross-functional teams.
- Proficiency in a systems programming language such as C++; the core engine is written in C++.
- Fluency in cloud environments such as AWS, GCP, or Azure and infrastructure-as-code tools such as Terraform.
- Experience with Linux systems, CI/CD pipelines, and modern observability stacks such as Prometheus and Grafana.
- Curiosity about database internals, storage engines, or low-latency query systems.
- Experience operating production clusters at scale, such as with Kubernetes or other orchestration systems.
- Rigorous approach to scalability, correctness, and reliability.
- Domain knowledge in databases, data systems, storage engines, indexing, or query processing is a plus but not required.
Benefits
- Equity and performance-related bonuses for eligible employees.
- Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
- Pre-tax accounts for health and dependent care expenses, parking, and transit.
- 401(k) retirement plan with employer match.
- Paid parental, medical, and caregiver leave.
- Paid time off, paid company holidays, office closures, and paid sick or safe time.
- Mental health and wellness support.
- Employer-paid basic life and disability coverage.
- Annual learning and development stipend.
- Daily meals in offices and eligible meal delivery credits.
- Relocation support for eligible employees.
More jobs at OpenAI
GRC Program Manager, Assurance Engineering & Control Systems
OpenAI · San Francisco, United States
USD 216,000-252,000 per year
Android Systems Engineer, Consumer Devices
OpenAI · San Francisco, United States
USD 216,000-342,000 per year
Senior Staff Software Engineer, Identity
OpenAI · Mountain View, United States, San Francisco, United States
USD 345,000-405,000 per year
Analytics Engineer, GTM
OpenAI · San Francisco, United States, New York City, United States
USD 220,000-335,000 per year
Product Designer, Payments
OpenAI · San Francisco, United States
USD 245,000-310,000 per year
Similar jobs
Senior Full-Stack Lead Engineer
Nvidia · Santa Clara, United States
USD 224,000-356,500 per year
Senior Software Engineer
SentinelOne · United States
USD 132,000-182,000 per year
Cloud Software Engineer - Observability Platform
ClickHouse · United States
USD 141,000-230,000 per year
Big Data Engineer With Java
ING · Katowice, Poland
PLN 156,000-336,000 per year
Infrastructure Security Engineer
SpaceXAI · Palo Alto, United States, Washington, United States, Austin, United States, New York City, United States
USD 100,000-258,000 per year
Senior Staff Network Automation Engineer
Nvidia · Santa Clara, United States
USD 208,000-333,500 per year
Senior Systems Software Engineer, Developer Productivity and Cloud Automation - GeForce NOW
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year
NCX Senior Engineer
Nvidia · Santa Clara, United States
USD 184,000-356,500 per year