Distributed Systems Engineer III

Mozn
Egypt
Workplace: RemoteFull timeFunction: IT Operations (Systems/Network Admin)Experience: 4-7 yearsSkills: []

Build and operate reliable, scalable cloud-native platforms for distributed, data-intensive workloads. Work hands-on with Kubernetes and GitOps tooling to deploy and troubleshoot production systems, and run Apache Kafka for high-throughput data flows. Support MySQL/PostgreSQL operations and reliability/DR across multi-tenant, multi-zone architectures. Automate lifecycle and observability using tools like Terraform, Prometheus, and Grafana, and help evolve the platform for AI and data workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Mozn
Mozn
1 day ago

Distributed Systems Engineer III

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Build and operate reliable, scalable cloud-native platforms for distributed, data-intensive workloads. Work hands-on with Kubernetes and GitOps tooling to deploy and troubleshoot production systems, and run Apache Kafka for high-throughput data flows. Support MySQL/PostgreSQL operations and reliability/DR across multi-tenant, multi-zone architectures. Automate lifecycle and observability using tools like Terraform, Prometheus, and Grafana, and help evolve the platform for AI and data workloads.
Location: Egypt
Workplace: Remote
Employment Type: Full time
Job Function: IT Operations (Systems/Network Admin)
Seniority: Mid level

Key Responsibilities

  • •Build, operate, and continuously improve production cloud-native platforms for distributed workloads.
  • •Work hands-on with Kubernetes operations, upgrades, node pools, workload lifecycle, and platform troubleshooting.
  • •Deploy and manage workloads using ArgoCD, Helm, GitOps, Terraform, and automation; simplify operational workflows.
  • •Operate and troubleshoot Apache Kafka in production, including topics/partitions, replication, consumer groups, retention, and failure recovery.
  • •Automate and improve reliability, observability, DR, and incident response using infrastructure and monitoring tools.

Pay and Benefits

Perks:Health Insurance

Key Requirements

  • •4–7 years of experience in platform/infrastructure engineering, distributed systems, SRE, backend engineering, or data infrastructure.
  • •Production hands-on experience with Apache Kafka (mandatory).
  • •Production hands-on experience with Kubernetes (mandatory).
  • •Strong experience with at least one relational database: MySQL or PostgreSQL (mandatory).
  • •Experience with infrastructure-as-code and GitOps (e.g., Terraform, ArgoCD/GitOps), plus distributed-systems fundamentals and reliability/incident troubleshooting.
Experience:4-7 yearsPlatform engineeringInfrastructure engineeringDistributed systemsSREData infrastructureCloud infrastructureDistributed data systemsAI/ML infrastructure
Tech Stack:KubernetesApache KafkaMySQLPostgreSQLArgoCDHelmGitOpsTerraformPythonBashGoJavaGCPOCIAWSAzureKafka ConnectDebeziumKafka StreamsPrometheus

Company Brief

Mozn
Builds enterprise AI products and data platforms that help organizations automate decision-making, improve risk management, and extract insights from large datasets. The company focuses on applied machine learning for regulated and data-intensive industries.
Industry: AI & Machine Learning
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Riyadh, Saudi Arabia
Founded: 2017
WebsiteLinkedIn