Mid/Senior Data Engineer

Yassir
Tunis
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringSkills: ["Communication","Root cause analysis","Collaboration","Technology research","Mentoring"]

Build and run a centralized data lake on GCP, integrating data across the enterprise. Develop and optimize Spark-powered batch and streaming ETL/ELT pipelines using GCP services such as Dataproc, Dataflow, Dataplex, Pub/Sub, BigQuery, and Cloud Storage. Implement data validation/quality checks, troubleshoot data issues with root-cause analysis, and collaborate with AI/ML, product, analysts, and business teams. Support GenAI initiatives and develop proofs of concept.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Yassir
Yassir
7 hours ago

Mid/Senior Data Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Build and run a centralized data lake on GCP, integrating data across the enterprise. Develop and optimize Spark-powered batch and streaming ETL/ELT pipelines using GCP services such as Dataproc, Dataflow, Dataplex, Pub/Sub, BigQuery, and Cloud Storage. Implement data validation/quality checks, troubleshoot data issues with root-cause analysis, and collaborate with AI/ML, product, analysts, and business teams. Support GenAI initiatives and develop proofs of concept.
Location: Tunis
Workplace: Hybrid
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Mid level

Key Responsibilities

  • •Integrate diverse enterprise data sources and build a centralized GCP-based data lake.
  • •Develop, maintain, and optimize Spark batch/streaming pipelines and ETL/ELT processes.
  • •Design data validation and quality checks to ensure accuracy, completeness, and consistency.
  • •Collaborate with AI/ML and other cross-functional teams to support analytical and machine learning use cases.
  • •Troubleshoot data issues, perform root-cause analysis, and support GenAI initiatives and PoCs.

Key Requirements

  • •Build and optimize Spark-powered batch and streaming data processing pipelines.
  • •Strong GCP data engineering experience (Dataproc, Dataflow, Dataplex, Pub/Sub, BigQuery, Cloud Storage).
  • •Hands-on programming skills in Scala and/or Python.
  • •Good SQL knowledge and experience with data governance, data warehousing, and data modeling.
  • •Experience with data quality frameworks (e.g., Great Expectations) and data validation checks.
Skills:CommunicationRoot cause analysisCollaborationTechnology researchMentoring
Tech Stack:GCPDataprocDataflowDataStreamDataplexPub/SubBigQueryCloud StorageSparkPySparkScalaPythonMongoDBGreat ExpectationsAirflowPrefectLuigiInfrastructure-as-CodeTerraformDocker

Company Brief

Yassir
Yassir is a North African super app offering ride-hailing, delivery, and fintech services. It connects riders, couriers, and merchants through a mobile platform focused on emerging markets across Africa.
Industry: Mobility Platforms
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Algiers, Algeria
Founded: 2017
Website