Site Reliability Engineer
Pythian
Hyderabad, India
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Collaboration","Automation","Reliability","Scalability"]Build and run resilient, high-performing infrastructure for large-scale distributed systems. You’ll operate and optimize Kubernetes clusters, Istio service mesh, and Linux-based environments, automate workflows with Go, Python, and Shell, and deliver observability with Prometheus, Grafana, and Loki. Collaborate with clients and AI/ML teams to ensure infrastructure readiness, while handling troubleshooting, on-call rotations, and postmortems to continuously improve reliability.

