Data Engineer Advisor

NTT
Chennai
Workplace: OnsiteFull timeFunction: Data Analytics & Business IntelligenceSkills: []

Design and provision the core Databricks data platform across multiple workspaces, establishing shared data structures and optimized storage using Delta Lake, Auto Loader, and Delta Live Tables. Configure cost-management strategies such as cluster sizing, autoscaling, and serverless SQL warehouses to balance performance and cloud spend. Apply deep Spark expertise (tuning, memory management, optimizations) and build reliable data infrastructure using Python (PySpark), SQL, and cloud/DevOps tooling.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NTT
NTT
5 days ago

Data Engineer Advisor

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Design and provision the core Databricks data platform across multiple workspaces, establishing shared data structures and optimized storage using Delta Lake, Auto Loader, and Delta Live Tables. Configure cost-management strategies such as cluster sizing, autoscaling, and serverless SQL warehouses to balance performance and cloud spend. Apply deep Spark expertise (tuning, memory management, optimizations) and build reliable data infrastructure using Python (PySpark), SQL, and cloud/DevOps tooling.
Location: Chennai
Workplace: Onsite
Employment Type: Full time
Job Function: Data Analytics & Business Intelligence

Key Responsibilities

  • •Architect and provision the core Databricks ecosystem across multi-workspace environments using Azure, AWS, or GCP.
  • •Establish shared data structures and optimized storage strategies using Delta Lake, Auto Loader, and Delta Live Tables (DLT).
  • •Implement cost-management strategies by configuring optimal cluster sizing, autoscaling, and serverless SQL warehouses to balance performance and cloud spend.
  • •Apply distributed systems and Apache Spark expertise to support data infrastructure performance and reliability.

Key Requirements

  • •Mastery of the Databricks ecosystem, including Unity Catalog, Delta Lake, Photon engine, Serverless Compute, Databricks Asset Bundles (DABs), and Delta Live Tables (DLT).
  • •Advanced proficiency in Python (PySpark), SQL, and shell scripting (Bash/PowerShell).
  • •Strong experience with Terraform, git workflows, and CI/CD tools such as Azure DevOps, GitHub Actions, or GitLab CI.
  • •Hands-on administration of data infrastructure on at least one cloud provider (AWS, Azure, or GCP).
  • •Strong understanding of Apache Spark core concepts, memory management, optimizations (Z-order, data skipping), and tuning.
Tech Stack:DatabricksUnity CatalogDelta LakePhoton engineServerless ComputeDatabricks Asset Bundles (DABs)Delta Live Tables (DLT)PythonPySparkSQLBashPowerShellTerraformGitAzure DevOpsGitHub ActionsGitLab CIAWSAzureGCP

Company Brief

NTT
Dimension Data, operating under NTT Ltd, provides managed IT services, cloud and data center solutions, networking, cybersecurity, and digital transformation services to enterprise customers worldwide.
Industry: Professional Services
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Valuation: Public Company (Market Cap in USD)
Headquarters: London, United Kingdom
Founded: 1983
WebsiteLinkedIn