Software Engineer, Data Integration

Aaru
New York
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 3+ yearsSkills: ["SQL","Python","Data integration","Data engineering","ETL/ELT","Entity resolution","Data quality"]

Build and maintain scalable data integration pipelines powering predictive simulations. In this role you’ll ingest and harmonize large multimodal datasets across APIs, flat files, cloud storage, and data warehouses, ensuring data quality and traceability for research and deployment. You’ll collaborate with engineering, research, and product teams to make integrated data readily usable for simulation ingestion at scale in a fast-moving, mission-driven environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Aaru
Aaru
3 months ago

Software Engineer, Data Integration

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 18 hours agoStatus: Live

Job Summary

Build and maintain scalable data integration pipelines powering predictive simulations. In this role you’ll ingest and harmonize large multimodal datasets across APIs, flat files, cloud storage, and data warehouses, ensuring data quality and traceability for research and deployment. You’ll collaborate with engineering, research, and product teams to make integrated data readily usable for simulation ingestion at scale in a fast-moving, mission-driven environment.
Location: New York
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Build and maintain scalable pipelines to ingest, clean, and integrate large multimodal datasets
  • •Own data ingestion across APIs, flat files, cloud storage, and data warehouses
  • •Design workflows for linkage, entity resolution, deduplication, and schema harmonization across imperfect or incongruent datasets
  • •Work with engineering, research, and deployment teams to make integrated data usable for simulation ingestion
  • •Establish and monitor data quality checks, validation logic, and documentation across datasets and pipelines

Pay and Benefits

Equity and Bonus:Equity
Perks:Health InsuranceVisionDentalEquityRelocationVisa Sponsorship

Key Requirements

  • •3+ years of experience in data integration, data engineering, ETL/ELT, or a similar role involving large-scale datasets
  • • Hands-on experience with messy, high-volume data (>100 TB) and building reliable systems at scale
  • •Fluency in SQL and Python, and familiarity with modern data infrastructure such as Snowflake, BigQuery, Databricks, or similar tools
  • •Strong judgment around data quality and ability to identify inconsistencies, edge cases, and integration risks
  • •Experience with entity resolution, data deduplication, and schema harmonization across diverse datasets
Experience:3+ yearsData integrationData engineeringETL/ELTDatasets
Skills:SQLPythonData integrationData engineeringETL/ELTEntity resolutionData quality
Tech Stack:SQLPythonSnowflakeBigQueryDatabricksAPIsCloud storage

Eligibility

Visa:Visa sponsorship
Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Aaru
Aaru builds multi-agent simulation software that models entire populations to predict behavior and future events, replacing traditional research with decision-ready forecasts for enterprises, governments, and agencies across industries.
Industry: AI & Machine Learning
Company Size: Small (11 to 50 employees)
Revenue: USD 1M to 5M
Growth: Early Stage Startup
Valuation: Unicorn (USD 1B+)
Funding: Series A
Headquarters: San Francisco, United States
Founded: 2024
WebsiteLinkedIn