Data Platform Modernization Engineer - Junior (Databricks)

NTT
Bengaluru
Workplace: OnsiteFull timeFunction: Healthcare (Clinical, Medical, Wellness)Experience: 3-6 yearsEducation: bachelorsSkills: ["Analytical","Problem-solving","Communication"]

Modernize data platforms by analyzing existing Informatica and AWS Glue ETL workflows and migrating them to a Databricks Lakehouse. Build Databricks notebooks, batch and incremental Delta Lake pipelines, and implement layered Bronze/Silver/Gold processing. Convert transformation logic while preserving business rules, perform unit testing and data reconciliation, troubleshoot pipeline issues, and support integration/UAT/production validation, monitoring, and cutover activities.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NTT
NTT
1 week ago

Data Platform Modernization Engineer - Junior (Databricks)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Modernize data platforms by analyzing existing Informatica and AWS Glue ETL workflows and migrating them to a Databricks Lakehouse. Build Databricks notebooks, batch and incremental Delta Lake pipelines, and implement layered Bronze/Silver/Gold processing. Convert transformation logic while preserving business rules, perform unit testing and data reconciliation, troubleshoot pipeline issues, and support integration/UAT/production validation, monitoring, and cutover activities.
Location: Bengaluru
Workplace: Onsite
Employment Type: Full time
Job Function: Healthcare (Clinical, Medical, Wellness)
Seniority: Mid level

Key Responsibilities

  • •Analyze existing Informatica workflows and AWS Glue jobs to understand source-to-target mappings and transformation logic.
  • •Develop Databricks notebooks and data pipelines to migrate ETL workloads to the Databricks Lakehouse.
  • •Implement Bronze/Silver/Gold layered processing and build batch and incremental pipelines using Delta Lake with watermark/control mechanisms.
  • •Convert Informatica transformations, mappings, lookups, and business rules into Databricks equivalents while preserving logic; reuse frameworks for logging, auditing, data quality, error handling, and notifications.
  • •Perform unit testing and source-to-target data reconciliation; troubleshoot data/SQL/PySpark/pipeline execution issues and support integration testing, regression testing, UAT, deployment, monitoring, and production cutover.

Key Requirements

  • •3–6 years of experience in data engineering, ETL development, or application/data platform development.
  • •Hands-on experience with Databricks and Apache Spark.
  • •Strong development skills in Python/PySpark and SQL, including Spark SQL/Databricks SQL.
  • •Experience building batch or incremental ETL/ELT pipelines and working with relational databases.
  • •Experience with unit testing, data validation, and troubleshooting, plus basic understanding of layered data architectures (Bronze/Silver/Gold).
Experience:3-6 yearsETLData engineering
Education:Bachelor's
Skills:AnalyticalProblem-solvingCommunication
Certifications:Databricks certification
Tech Stack:DatabricksApache SparkPythonPySparkSpark SQLDatabricks SQLDelta LakeDatabricks JobsDatabricks WorkflowsLakeflow Declarative PipelinesDelta Live TablesUnity CatalogDatabricks Asset BundlesCI/CDAWS S3AWS GlueAmazon RedshiftAuto LoaderCDCSCD Type 1

Company Brief

NTT
Dimension Data, operating under NTT Ltd, provides managed IT services, cloud and data center solutions, networking, cybersecurity, and digital transformation services to enterprise customers worldwide.
Industry: Professional Services
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Valuation: Public Company (Market Cap in USD)
Headquarters: London, United Kingdom
Founded: 1983
WebsiteLinkedIn