Data Engineer, People Innovation Labs
Core
Build data-intensive systems and pipelines to power internal people products (like OpenHouse) and enable People Analytics at OpenAI.
Role type
Senior Data Engineer
Builds
Internal people products, data pipelines, canonical datasets for people metrics
Domain
Internal HR/Tech, Data Engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Python, Scala, Java, Databricks, Snowflake, ETL schedulers (Airflow, Dagster, Prefect), Spark, Hadoop, Flink, distributed storage (HDFS, S3)
Preferred skills
null
Technologies
Databricks, Snowflake, Fivetran, Airflow, Dagster, Prefect, Spark, Hadoop, Flink, HDFS, S3
Responsibilities
Design and manage people data pipelines integrated into Databricks; Develop canonical datasets for people and product metrics; Collaborate with Data Platform, Data Science, and People Analytics teams; Implement fault-tolerant data ingestion and processing systems; Participate in data architecture decisions; Ensure data security, integrity, and compliance
Seniority
Senior, hands-on IC