Senior Site Reliability Engineer

Fiserv
Sunnyvale
Workplace: OnsiteFull timeUSD 160,000 - 240,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Troubleshooting","Problem-solving","Communication"]

Join the global reliability team in Sunnyvale to operate financial platforms at scale. Build automation to remove manual operational work, enhance monitoring, logging, and alerting, and lead incident response with post-incident RCA. Define SLIs/SLOs and manage error budgets, forecast capacity for performance and cost efficiency, and troubleshoot production issues. Work with an international team to harden cloud-native platforms across GCP/GKE.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Fiserv
Fiserv
2 days ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Join the global reliability team in Sunnyvale to operate financial platforms at scale. Build automation to remove manual operational work, enhance monitoring, logging, and alerting, and lead incident response with post-incident RCA. Define SLIs/SLOs and manage error budgets, forecast capacity for performance and cost efficiency, and troubleshoot production issues. Work with an international team to harden cloud-native platforms across GCP/GKE.
Location: Sunnyvale
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design, build, and maintain automation to eliminate manual repetitive operational tasks, including runbooks, deployment pipelines, and remediation scripts.
  • •Operate and enhance monitoring, logging, and alerting systems to ensure observability across services.
  • •Participate in on-call rotations and lead incident response, running and documenting post-incident RCA and follow-up actions.
  • •Collaborate with stakeholders to define SLIs and SLOs, manage error budgets, and translate reliability goals into measurable actions.
  • •Troubleshoot production issues through root-cause analysis and implement durable fixes while driving continuous improvement and platform hardening.

Pay and Benefits

Salary: USD 160,000 - 240,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Solid practical experience in site reliability, operations, or DevOps at a mid-to-senior level.
  • •Strong shell scripting skills and a foundation in programming concepts.
  • •Hands-on cloud experience, specifically Google Cloud Platform (GCP) and GKE.
  • •Proven experience with containerisation and orchestration (Kubernetes).
  • •Familiarity with monitoring and observability tooling such as Prometheus, Grafana, and Datadog.
Experience:FintechPaymentsCloud
Skills:TroubleshootingProblem-solvingCommunication
Tech Stack:Google Cloud Platform (GCP)GKEKubernetesTerraformAnsiblePuppetPrometheusGrafanaDatadogHTTP(s)HAProxyGitHubGitHub ActionsInfrastructure as CodeContainerisationObservability

Company Brief

Fiserv
Provides payments, processing services, risk management, and core banking technology to financial institutions, merchants, and businesses worldwide, enabling digital payments, account processing, and financial services integration.
Industry: Fintech Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Brookfield, United States
Founded: 1984
WebsiteLinkedIn