Data Engineer, Analytics Data Products

The New York Times
New York
Workplace: OnsiteFull timeUSD 110,000 - 130,000 annuallyFunction: Data Analytics & Business IntelligenceExperience: 2+ yearsSkills: ["SQL","Python","Data modeling","Spark","Dbt","PySpark","Git","Airflow","BigQuery","Cloud","ETL","ELT"]

Data Engineer for analytics data products, owning end-to-end ELT/ETL pipelines, data modeling, and platform optimization across GCP and AWS. You’ll use dbt, PySpark, and Spark to build scalable data products, collaborate with cross-functional teams, and ensure high-quality, observable data for analytics tools in a hybrid, remote-friendly NY-based engineering environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
The New York Times
The New York Times
7 months ago

Data Engineer, Analytics Data Products

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 hours agoStatus: Live

Job Summary

Data Engineer for analytics data products, owning end-to-end ELT/ETL pipelines, data modeling, and platform optimization across GCP and AWS. You’ll use dbt, PySpark, and Spark to build scalable data products, collaborate with cross-functional teams, and ensure high-quality, observable data for analytics tools in a hybrid, remote-friendly NY-based engineering environment.
Location: New York
Workplace: Onsite
Employment Type: Full time
Job Function: Data Analytics & Business Intelligence

Key Responsibilities

  • •Design, model, and implement complex ELT/ETL pipelines for the cleansed and curated data layers in the medallion architecture, taking full ownership of the data product's structure, partitioning, documentation, and performance characteristics.
  • •Develop advanced data transformations using dbt (data build tool) for relational data modeling and PySpark for large-scale data processing within the Lakehouse, ensuring outputs meet strict Service Level Agreements and quality standards.
  • •Collaborate across teams to define requirements and translate them into robust and scalable data models suitable for analytic consumption.
  • •Manage the physical data storage across both GCP and AWS, selecting optimal file formats and designing efficient partitioning and clustering strategies.
  • •Administer and tune Spark compute resources (e.g., Dataproc, EMR, or managed services) to optimize job execution time and cost.

Pay and Benefits

Salary: USD 110,000 - 130,000 annually
Perks:Health InsuranceDentalVision401kPaid Leave

Key Requirements

  • •2+ years of hands-on experience in a Data Engineering, Data Warehousing, Analytics Engineering or equivalent role.
  • •Proficiency in SQL and experience with complex, production-level data modeling (dimensional modeling, Kimball, OBT, or Data Vault)
  • •Demonstrated experience designing, developing, and deploying end-to-end data products through the full Software Development Lifecycle
  • •Experience with a Cloud Data Warehouse, like BigQuery
  • •Proficiency in Python for scripting and data manipulation, including knowledge of PySpark or other Spark APIs
Experience:2+ yearsData engineeringData warehousingAnalytics
Skills:SQLPythonData modelingSparkDbtPySparkGitAirflowBigQueryCloudETLELT
Languages:English
Tech Stack:SQLDbtPySparkSparkBigQueryPythonGCPAWSDataprocEMRAirflowCloud ComposerPrefectGitTerraformIcebergDelta LakeHexRBAC

Company Brief

The New York Times
Operates The New York Times, a global news organization delivering journalism across print and digital platforms, plus specialty sites like Wirecutter. Produces reporting, opinion, multimedia, and subscription-driven content covering national and international affairs.
Industry: News, Journalism & Broadcasting
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 1851
WebsiteLinkedIn