Site Reliability Consultant

Pythian
Canada
Workplace: RemoteFull timeCAD 90,000 - 100,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Root cause analysis","Troubleshooting","Collaboration","Technology leadership","Fast learning"]

Build and scale a next-generation Site Reliability Engineering practice for customer infrastructure. You will operate and improve production systems for availability, visibility, and resiliency, perform root-cause analysis for incidents, and automate repetitive operational work. Partner with clients on technology roadmaps and work through on-call escalations. The role focuses on cloud reliability across Google Cloud and AWS, automation, and containerized microservices environments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Pythian
Pythian
3 days ago

Site Reliability Consultant

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Build and scale a next-generation Site Reliability Engineering practice for customer infrastructure. You will operate and improve production systems for availability, visibility, and resiliency, perform root-cause analysis for incidents, and automate repetitive operational work. Partner with clients on technology roadmaps and work through on-call escalations. The role focuses on cloud reliability across Google Cloud and AWS, automation, and containerized microservices environments.
Location: Canada
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Operate, maintain, and administer solutions that improve customer infrastructure operational efficiency, availability, and visibility.
  • •Plan maintenance activity and create design documentation and standard procedures.
  • •Provide root-cause analysis reports for outages/incidents using ITIL problem management.
  • •Monitor client infrastructure, identify resiliency improvements, reduce incident occurrence, and automate repetitive tasks.
  • •Participate in an on-call rotation in an escalation capacity and lead client discussions on technology roadmaps.

Pay and Benefits

Salary: CAD 90,000 - 100,000 annually
Perks:Remote WorkLearning BudgetWellness StipendPaid LeavePaid Sick

Key Requirements

  • •Experience working with Google Cloud (necessary) and AWS Clouds, including infrastructure-as-code deployment.
  • •Scripting and automation of administrative tasks using Python and Scala.
  • •Solid understanding of microservices architecture and container technologies (Kubernetes is a must; Docker, lxc, etc.).
  • •Clear understanding of software development lifecycles and best practices from an infrastructure perspective.
  • •Comprehensive Linux systems administration and troubleshooting, including cloud migration and network troubleshooting (TCP/IP, DNS, NTP, DHCP, SMTP, etc.).
Experience:CloudDevOps
Skills:Root cause analysisTroubleshootingCollaborationTechnology leadershipFast learning
Tech Stack:Google CloudAWSCloudFormationTerraformOpsworksPythonScalaMicroservicesKubernetesDockerLxcLinuxTCP/IPDNSNTPDHCPSMTPHypervisorVMwareHyper-V

Company Brief

Pythian
Provides data, cloud, and managed services to help organizations migrate, operate, and optimize analytics, databases, and cloud infrastructure across hybrid and multi-cloud environments.
Industry: Consulting
Company Size: Large (251 to 1,000 employees)
Growth: Established Company
Headquarters: Ottawa, Canada
Founded: 1997
WebsiteLinkedIn