Data Engineer (SQL, PySpark, Snowflake, DBT, Redshift) - Associate

KKR
Gurugram
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 5-8 yearsSkills: ["Communication","Cross-functional collaboration","Proactive ownership","Mentorship","Problem-solving"]

Design, build, and operate production-grade data platforms on AWS to support analytics and business decisions across Insurance Systems. Own the full lifecycle of end-to-end data pipelines—from ingestion and orchestration to transformation, storage, and governance—using a modern lakehouse stack (Apache Iceberg, AWS Glue, Snowflake). Collaborate with data scientists and platform teams to deliver reliable, well-governed, and cost-efficient data products, including data quality, lineage, and security.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
KKR
KKR
2 days ago

Data Engineer (SQL, PySpark, Snowflake, DBT, Redshift) - Associate

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Design, build, and operate production-grade data platforms on AWS to support analytics and business decisions across Insurance Systems. Own the full lifecycle of end-to-end data pipelines—from ingestion and orchestration to transformation, storage, and governance—using a modern lakehouse stack (Apache Iceberg, AWS Glue, Snowflake). Collaborate with data scientists and platform teams to deliver reliable, well-governed, and cost-efficient data products, including data quality, lineage, and security.
Location: Gurugram
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Design, build, and maintain scalable, fault-tolerant data pipelines and ETL/ELT processes for structured, semi-structured, and unstructured data.
  • •Own end-to-end pipeline orchestration including scheduling, dependency management, retries, SLAs, and observability.
  • •Build and manage lakehouse data assets on Apache Iceberg, including partitioning, schema evolution, time-travel, and table maintenance.
  • •Develop and operate large-scale data processing jobs using Apache Spark (including AWS Glue Spark jobs) and model/load/optimize data in Snowflake.
  • •Ensure data quality, integrity, lineage, and security; integrate data from APIs, relational databases, streaming feeds, and third-party tools while monitoring and tuning pipeline performance and cost.

Key Requirements

  • •5 to 8 years of hands-on data engineering experience building production data pipelines.
  • •Strong proficiency in Python for data engineering and automation.
  • •Advanced SQL and strong experience with relational databases (e.g., PostgreSQL, MySQL).
  • •Experience operating a workflow orchestration tool in production (e.g., Apache Airflow, Dagster, or equivalent).
  • •Hands-on experience with Apache Spark and lakehouse storage using Apache Iceberg (or an equivalent open table format).
Experience:5-8 years
Skills:CommunicationCross-functional collaborationProactive ownershipMentorshipProblem-solving
Languages:English
Tech Stack:AWSS3AWS GlueEMRLambdaAthenaKinesisRedshiftIAMApache IcebergApache SparkSnowflakeApache AirflowDagsterDockerKubernetesAmazon EKSTerraformAWS CloudFormationCDK

Company Brief

KKR
Global investment firm providing alternative asset management and capital markets services across private equity, credit, real assets, and hedge funds, serving institutional and private clients worldwide.
Industry: Asset Management
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 1976
WebsiteLinkedIn