OpenAI is seeking a Data Engineer based in California to lead in building data pipelines essential for analyses, safety systems, and product growth. This role involves close collaboration with teams across infrastructure, data science, and product. Candidates should have 3+ years in data engineering and experience with technologies like Python, Hadoop, and ETL schedulers. The position includes relocation assistance, embodying a commitment to harnessing AI for the benefit of all humanity.
Qualifications
3+ years of experience as a data engineer and 8+ years in software engineering.
Proficient in at least one programming language used in Data Engineering.
Experience with distributed processing technologies.
Responsibilities
Design, build, and manage data pipelines to integrate user event data.
Develop datasets to track product metrics including user growth.
Collaborate with various teams to understand data needs.
Skills
Data Engineering
Python
Scala
Java
Distributed Processing Technologies
ETL Scheduling
Spark
Tools
Hadoop
Flink
HDFS
S3
Airflow
Dagster
Prefect
Job description
OpenAI is seeking a Data Engineer based in California to lead in building data pipelines essential for analyses, safety systems, and product growth. This role involves close collaboration with teams across infrastructure, data science, and product. Candidates should have 3+ years in data engineering and experience with technologies like Python, Hadoop, and ETL schedulers. The position includes relocation assistance, embodying a commitment to harnessing AI for the benefit of all humanity.