Data Engineer (Python AND Kafka AND Hadoop OR HDFS OR Hive OR Snowflake AND apache AND iceberg)
Dimension Data
Anywhere
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 3-5 yearsEducation: bachelorsSkills: ["Integrity","Collaboration","Clear communication","Stakeholder management","Ownership"]Build and migrate end-to-end datastore pipelines from an on-prem DataLake to an AWS-hosted Lakehouse for a high-visibility program. Refactor extraction logic and scheduling, transfer datasets with data integrity, and convert legacy SQL/Spark consumption patterns for Snowflake and Iceberg. Own reconciliation and quality validation to ensure migrated assets match production usage, while partnering with data owners through handoff and sign-off across global teams.

