Site Reliability Engineer
Riyadh
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsSkills: ["Ownership","Incident management","Methodical troubleshooting","Clear communication","Automation mindset"]Own the reliability of cloud infrastructure and ensure the platform stays stable and scalable as real-time customer data volumes grow. Design fault-tolerant systems, eliminate single points of failure, and manage workloads across AWS, GCP, or Azure using Terraform and Kubernetes. Build observability with tools like Prometheus, Grafana, Datadog, and ELK, drive incident response and root-cause analysis, and automate repetitive operational work to improve deployment reliability.
Loading
Loading job details...
Preparing the role view and application actions.

