Site Reliability Engineer

Pythian
Argentina, Brazil, Uruguay
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Collaboration","Automation","Scalability","Reliability"]

Build and operate resilient, high-performing infrastructure as part of Pythian’s next-generation SRE team. You’ll run and optimize Kubernetes clusters with Istio and Linux systems, automate workflows with Go, Python, and Shell, and deliver monitoring/observability using Prometheus, Grafana, and Loki. Tackle challenging networking, storage, and performance issues, support AI/ML readiness, and participate in on-call rotations and postmortems to improve reliability.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Pythian
Pythian
17 hours ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 hours agoStatus: Live

Job Summary

Build and operate resilient, high-performing infrastructure as part of Pythian’s next-generation SRE team. You’ll run and optimize Kubernetes clusters with Istio and Linux systems, automate workflows with Go, Python, and Shell, and deliver monitoring/observability using Prometheus, Grafana, and Loki. Tackle challenging networking, storage, and performance issues, support AI/ML readiness, and participate in on-call rotations and postmortems to improve reliability.
Location: Argentina, Brazil, Uruguay
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Operate and optimize Kubernetes clusters, Istio service mesh, and Linux-based systems.
  • •Automate workflows using Go, Python, and Shell scripting.
  • •Build monitoring and observability solutions with Prometheus, Grafana, and Loki.
  • •Troubleshoot complex networking, storage, and system performance issues.
  • •Partner with AI/ML teams and participate in on-call rotations and postmortems to improve resilience.

Pay and Benefits

Perks:Remote WorkLearning BudgetGym MembershipPaid LeaveWellness Stipend

Key Requirements

  • •Experience with Google Cloud plus infrastructure-as-code (IaC) tools such as Terraform.
  • •Strong knowledge of microservices and containers, including Kubernetes and Docker, plus networking fundamentals.
  • •Hands-on automation experience using Go, Python, and Shell scripting.
  • •Experience building monitoring and observability solutions with Prometheus, Grafana, and Loki.
  • •Hands-on Linux administration and experience with PKI and service mesh (e.g., Istio).
Experience:SRECloudKubernetesMicroservicesObservabilityAI/ML
Skills:Problem-solvingCollaborationAutomationScalabilityReliability
Tech Stack:Google CloudAWSKubernetesIstioLinuxGoPythonShell scriptingTerraformPrometheusGrafanaLokiPKIService meshMicroservicesDockerNetworkingStorageAI/MLDistributed systems

Company Brief

Pythian
Provides data, cloud, and managed services to help organizations migrate, operate, and optimize analytics, databases, and cloud infrastructure across hybrid and multi-cloud environments.
Industry: Consulting
Company Size: Large (251 to 1,000 employees)
Growth: Established Company
Headquarters: Ottawa, Canada
Founded: 1997
WebsiteLinkedIn