Senior AI Data Pipeline Engineer (Autonomous Driving)

42dot
Sunnyvale, San Francisco
Workplace: HybridFull timeUSD 133,000 - 254,000 annuallyFunction: Data Science & Machine LearningExperience: 7+ yearsEducation: bachelorsSkills: ["Leadership","Communication"]

Build large-scale, high-reliability data pipelines for autonomous driving datasets, from extracting millions of raw scenes to enabling data labeling and performance optimization. Develop an autonomous driving data SDK for scene search and dataset preparation, and help establish a data lakehouse for sensor, calibration, and annotation data. Partner with ML algorithm, ML application, and cloud infrastructure teams to align data platform components with the overall autonomous driving architecture.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
42dot
42dot
4 months ago

Senior AI Data Pipeline Engineer (Autonomous Driving)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Build large-scale, high-reliability data pipelines for autonomous driving datasets, from extracting millions of raw scenes to enabling data labeling and performance optimization. Develop an autonomous driving data SDK for scene search and dataset preparation, and help establish a data lakehouse for sensor, calibration, and annotation data. Partner with ML algorithm, ML application, and cloud infrastructure teams to align data platform components with the overall autonomous driving architecture.
Location: Sunnyvale, San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Develop high-scale, reliable data extraction pipelines to convert millions of raw scenes into high-value scene data.
  • •Build data labeling pipelines to perform auto-labeling inferences for autonomous driving algorithms.
  • •Develop an autonomous driving data SDK for scene data search and dataset preparation/loading.
  • •Build and maintain a data lakehouse for autonomous driving datasets (sensor, calibration, and annotation data).
  • •Diagnose pipeline performance bottlenecks and bootstrap/maintain infrastructure for data platform components; collaborate with ML and cloud teams on system architecture alignment.

Pay and Benefits

Salary: USD 133,000 - 254,000 annually

Key Requirements

  • •Bachelor's degree or higher in Computer Science, Engineering, Robotics, or a similar technical field.
  • •Minimum of 7 years of experience in Data Engineering, DataOps, or ML Platform roles.
  • •Proficient in Python with solid experience in Python SDK development.
  • •Hands-on experience orchestrating data pipeline jobs with Databricks Workflows or Apache Airflow and integrating pipelines with ML models.
  • •Experience with databases (e.g., MongoDB, PostgreSQL) and data architectures such as Data Warehouse (Hive) or Lakehouse (Delta Lake), plus Apache Spark or other big data engines.
Experience:7+ yearsData engineeringDataopsMl platformsAutonomous driving
Education:Bachelor's
Skills:LeadershipCommunication
Tech Stack:PythonDatabricks WorkflowsApache AirflowMongoDBPostgreSQLHiveDelta LakeApache SparkPyTorchTensorFlow

Company Brief

42dot
Develops autonomous driving and mobility software platforms, including AI-based perception, mapping, routing, and connected-vehicle technologies. It works on next-generation transportation systems and self-driving vehicle capabilities for automotive applications.
Industry: Autonomous Vehicles
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Seoul, South Korea
Founded: 2019
WebsiteLinkedIn