Site Reliability Consultant
Pythian
Ottawa, Vancouver
Workplace: HybridFull timeCAD 90,000 - 100,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Automation mindset","Scalability","Reliability","Collaboration"]Design, deploy, and operate large-scale distributed infrastructure as part of a next-generation SRE team. You’ll operate and optimize Kubernetes clusters with Istio and Linux systems, automate workflows, and build monitoring/observability using Prometheus, Grafana, and Loki. Tackle complex networking, storage, and performance issues, partner with AI/ML teams to keep training and data pipelines ready, and improve resilience through on-call and postmortems.

