Data Engineer, People Innovation Labs

OpenAI
San Francisco
Workplace: HybridFull timeUSD 293,000 - 325,000 annuallyFunction: Laboratory & Clinical OperationsExperience: 3+ yearsSkills: ["Python","Scala","Java","Data warehousing","Databricks","Snowflake","Airflow","Dagster","Prefect","Fivetran","Spark","Hadoop","Flink","Hdfs","S3"]

Design and build data-intensive pipelines powering OpenAI’s People Innovation Labs products (e.g., OpenHouse). Develop canonical datasets, collaborate with Data Platform/Data Science/People Analytics teams, and ensure robust, secure data ingestion and processing in a Databricks/Snowflake data warehouse. Lead data engineering decisions as the team’s technical expert while enabling organizational analytics and metrics.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
4 months ago

Data Engineer, People Innovation Labs

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 20 minutes agoStatus: Live

Job Summary

Design and build data-intensive pipelines powering OpenAI’s People Innovation Labs products (e.g., OpenHouse). Develop canonical datasets, collaborate with Data Platform/Data Science/People Analytics teams, and ensure robust, secure data ingestion and processing in a Databricks/Snowflake data warehouse. Lead data engineering decisions as the team’s technical expert while enabling organizational analytics and metrics.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Laboratory & Clinical Operations

Key Responsibilities

  • •Design, build and manage people data pipelines, ensuring all data is seamlessly integrated into our Databricks warehouse.
  • •Develop canonical datasets to track key people metrics and product metrics for the organization.
  • •Collaborate with Data Platform, Data Science, People Analytics, and Compensation/Equity teams to understand data needs and provide solutions.
  • •Implement robust and fault-tolerant systems for data ingestion and processing.
  • •Participate in data architecture and engineering decisions as the primary data engineering expert on the team.

Pay and Benefits

Salary: USD 293,000 - 325,000 annually
Equity and Bonus:Equity

Key Requirements

  • •3+ years of experience as a data engineer and 8+ years of any software engineering experience (including data engineering)
  • •Proficiency in Python, Scala, or Java
  • •Experience with data warehousing technologies such as Databricks and Snowflake, and ETL schedulers such as Fivetran, Airflow, Dagster, Prefect, or similar
  • •Experience with distributed processing technologies and frameworks such as Spark, Hadoop, Flink and distributed storage systems (e.g., HDFS, S3)
  • •Design, build and manage data pipelines and ensure data is integrated into Databricks warehouse
Experience:3+ yearsData engineeringAi
Skills:PythonScalaJavaData warehousingDatabricksSnowflakeAirflowDagsterPrefectFivetranSparkHadoopFlinkHdfsS3
Tech Stack:PythonScalaJavaDatabricksSnowflakeFivetranAirflowDagsterPrefectSparkHadoopFlinkHDFSS3

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor