Senior Data Engineer, Bioinformatics, Cheminformatics, Materials

Lila Sciences
San Francisco
Workplace: OnsiteFull timeUSD 144,000 - 240,000 annuallyFunction: Solutions Engineering & Sales EngineeringExperience: 2-6 yearsSkills: ["Data validation","Data quality","Problem-solving","Automation","Observability"]

Build ETL pipelines and data models for a scientific data platform, turning raw lab instrument outputs into validated, analysis-ready datasets. Model heterogeneous bio, chemistry, and materials data, implement data quality and schema-evolution checks, and develop reusable analysis functions to accelerate scientific and AI workflows. Improve automation, observability, and reliability across instrument-to-result data flows using strong Python, SQL, and workflow orchestration.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Lila Sciences
Lila Sciences
5 days ago

Senior Data Engineer, Bioinformatics, Cheminformatics, Materials

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 hours agoStatus: Live

Job Summary

Build ETL pipelines and data models for a scientific data platform, turning raw lab instrument outputs into validated, analysis-ready datasets. Model heterogeneous bio, chemistry, and materials data, implement data quality and schema-evolution checks, and develop reusable analysis functions to accelerate scientific and AI workflows. Improve automation, observability, and reliability across instrument-to-result data flows using strong Python, SQL, and workflow orchestration.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Mid level

Key Responsibilities

  • •Design pipelines that turn raw lab output into analysis-ready scientific data.
  • •Model heterogeneous data from bio, chemistry, and materials instruments.
  • •Build validation checks, schema-evolution gates, and data quality workflows.
  • •Develop reusable analysis functions and reusable data transformations for scientific and AI research workflows.
  • •Improve automation and observability across instrument-to-result data flows.

Pay and Benefits

Salary: USD 144,000 - 240,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVisionLife InsuranceDisability InsurancePaid ParentalLearning BudgetMeal Allowance

Key Requirements

  • •2–6 years of experience in data engineering, bioinformatics, cheminformatics, or computational science.
  • •Strong Python skills, including typed, tested, production-quality code.
  • •Strong SQL skills, especially with Postgres or similar relational databases.
  • •Experience building ETL pipelines, data models, and reusable data transformations.
  • •Workflow orchestration experience, ideally Flyte, Airflow, Prefect, Dagster, or Nextflow.
Experience:2-6 yearsBioinformaticsCheminformaticsComputational science
Skills:Data validationData qualityProblem-solvingAutomationObservability
Languages:English
Tech Stack:PythonSQLPostgresETLPandasNumPyFlyteAirflowPrefectDagsterNextflowParquetIcebergDuckDBPolarsIbisNATSKafkaLIMSELN

Company Brief

Lila Sciences
Develops AI-driven platforms to accelerate drug discovery and biological research by integrating machine learning with chemical and biological data to predict molecular properties, streamline candidate selection, and enable faster therapeutic development.
Industry: Biotech
Company Size: Small (11 to 50 employees)
Growth: Early Stage Startup
Funding: Seed
WebsiteLinkedIn