Data Engineer

at OpenAI
USD 230,000-385,000 per year
SENIOR
✅ On-site
✅ Relocation

Tech Stack

AI Airflow @ 6 ChatGPT Compliance Dagster @ 6 Data Engineering @ 7 Data Pipelines Data Science @ 4 ETL @ 6 Flink @ 4 Hadoop @ 4 Java @ 6 Marketing @ 4 Python @ 6 Scala @ 6 Security Spark @ 4

Details

The Applied team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. The team seeks to learn from deployment and distribute the benefits of AI while ensuring that the technology is used responsibly and safely.

The role focuses on building data pipelines and core tables that power analyses, safety systems, business decisions, product growth, and efforts to prevent bad actors. The Data Engineer will collaborate closely with the researchers behind ChatGPT and help train new models for users.

Responsibilities

  • Design, build, and manage data pipelines, ensuring that user event data is seamlessly integrated into the data warehouse.
  • Develop canonical datasets to track key product metrics, including user growth, engagement, and revenue.
  • Collaborate with Infrastructure, Data Science, Product, Marketing, Finance, and Research teams to understand data needs and provide solutions.
  • Implement robust and fault-tolerant systems for data ingestion and processing.
  • Participate in data architecture and engineering decisions.
  • Ensure the security, integrity, and compliance of data according to industry and company standards.

Requirements

  • 3+ years of experience as a data engineer and 8+ years of software engineering experience, including data engineering.
  • Proficiency in at least one programming language commonly used in data engineering, such as Python, Scala, or Java.
  • Experience with distributed processing technologies and frameworks such as Hadoop and Flink.
  • Experience with distributed storage systems such as HDFS and Amazon S3.
  • Expertise with ETL schedulers such as Airflow, Dagster, Prefect, or similar frameworks.
  • Solid understanding of Apache Spark, including the ability to write, debug, and optimize Spark code.

Work Location

This role is exclusively based at the San Francisco headquarters.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. OpenAI is an equal opportunity employer and does not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristics. Background checks are administered in accordance with applicable law. Reasonable accommodations are available to applicants with disabilities.

Benefits

  • Base salary range of $230,000–$385,000 per year.
  • Equity, performance-related bonuses for eligible employees, and benefits.
  • Medical, dental, and vision insurance, with employer contributions to Health Savings Accounts.
  • Pre-tax accounts for health and dependent care expenses, as well as commuter expenses.
  • 401(k) retirement plan with employer match.
  • Paid parental, medical, and caregiver leave.
  • Paid time off, company holidays, office closures, and paid sick or safe time as required by applicable law.
  • Mental health and wellness support.
  • Employer-paid basic life and disability coverage.
  • Annual learning and development stipend.
  • Daily office meals and eligible meal delivery credits.
  • Relocation support for eligible employees.

More jobs at OpenAI

Similar jobs