Site Reliability Engineering Manager

CAE
Krakow
Workplace: OnsiteFull timePLN 23,700 - 32,000 monthlyFunction: DevOps, Cloud & InfrastructureSkills: ["Leadership","Analytical mindset","Communication","Cross-functional collaboration","Automation mindset"]

Lead a reliability engineering team to ensure critical production platforms are highly available, measurable, resilient, and scalable. Own reliability strategy and operating model, define SLIs/SLOs and error budgets, and drive incident management and postmortems. Establish observability across databases, services, and infrastructure, automate operations with Infrastructure as Code and self-healing, and partner across engineering, security, and product to improve reliability end to end.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
CAE
CAE
1 month ago

Site Reliability Engineering Manager

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 days agoStatus: Live

Job Summary

Lead a reliability engineering team to ensure critical production platforms are highly available, measurable, resilient, and scalable. Own reliability strategy and operating model, define SLIs/SLOs and error budgets, and drive incident management and postmortems. Establish observability across databases, services, and infrastructure, automate operations with Infrastructure as Code and self-healing, and partner across engineering, security, and product to improve reliability end to end.
Location: Krakow
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Manager level

Key Responsibilities

  • •Lead and develop a team of reliability engineers across database and platform domains.
  • •Define reliability strategy, operating model, and roadmap for critical production services.
  • •Own availability, performance, scalability, and resilience for production databases and dependent platform services.
  • •Define and enforce SLIs, SLOs, and error budgets; drive incident management, postmortems, and remediation.
  • •Establish observability strategy and dashboards/alerting, and drive automation/self-healing and performance tuning/capacity planning.

Pay and Benefits

Salary: PLN 23,700 - 32,000 monthly
Perks:MultisportHealth InsuranceTravel AllowanceLife InsuranceEmployee AssistanceEquityReferral ProgramLearning BudgetBereavement LeaveVolunteer LeaveMaternal AllowancePensionTax-deductible Costs

Key Requirements

  • •7+ years of experience in Site Reliability Engineering, Database Engineering, or related production engineering roles.
  • •2+ years of leadership experience managing engineers or technical teams.
  • •Strong expertise with relational databases such as PostgreSQL, Oracle, or SQL Server.
  • •Experience operating reliable services in cloud or hybrid environments, including AWS-managed or self-managed platforms.
  • •Strong background in automation, observability, incident response, performance tuning, and high availability/disaster recovery practices.
Experience:Production engineeringCloudHybrid
Skills:LeadershipAnalytical mindsetCommunicationCross-functional collaborationAutomation mindset
Tech Stack:PostgreSQLOracleSQL ServerAWSKubernetesInfrastructure as Code

Company Brief

CAE
Provides simulation technologies, integrated training services, and flight training for civil aviation, defense, and healthcare customers worldwide, delivering pilot training, mission rehearsal, and simulation solutions.
Industry: Aviation Services
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Montreal, Canada
Founded: 1947
WebsiteLinkedIn