Staff Site Reliability Engineer

Fingerprint
United States
Workplace: RemoteFull timeUSD 177,000 - 240,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsSkills: ["Incident leadership","Leading through influence","Written communication","Teaching/coaching","Pragmatism"]

Be the first dedicated SRE to set reliability standards across Fingerprint’s engineering groups. Define SLIs/SLOs and error budgets, strengthen incident response and postmortems, and improve alert quality and production readiness through deliberate failure testing. Embed with teams to codify SRE practices that persist, stay hands-on during incidents, and lead AI-assisted reliability adoption across runbooks, observability, and tooling. Operate without direct reports, reporting to the VP of Engineering.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Fingerprint
Fingerprint
2 days ago

Staff Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Be the first dedicated SRE to set reliability standards across Fingerprint’s engineering groups. Define SLIs/SLOs and error budgets, strengthen incident response and postmortems, and improve alert quality and production readiness through deliberate failure testing. Embed with teams to codify SRE practices that persist, stay hands-on during incidents, and lead AI-assisted reliability adoption across runbooks, observability, and tooling. Operate without direct reports, reporting to the VP of Engineering.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Define SLIs and SLOs for critical request paths with team owners, making metrics visible and tied to decisions.
  • •Introduce error budgets to balance reliability investment against feature work and coach EMs/staff engineers on using them.
  • •Strengthen the end-to-end incident lifecycle: detection, response, communication, postmortem quality, and completion of follow-ups.
  • •Improve alerting and escalation design, working with Cloud Platform on shared tooling and closing the “customers find out before we do” gap.
  • •Build the SRE mindset by embedding with teams, codifying practices (production readiness, on-call standards, runbook quality, change safety), and testing for failure deliberately (game days/chaos).

Pay and Benefits

Salary: USD 177,000 - 240,000 annually

Key Requirements

  • •10+ years of engineering experience with 3+ years as an SRE, production engineer, or reliability-focused Staff engineer, owning reliability for a platform across multiple teams.
  • •Practical expertise designing SLIs/SLOs and using error budgets, including driving product/team adoption.
  • •Strong incident leadership for high-severity, customer-facing incidents and improved organizational learning via postmortems.
  • •Hands-on depth in distributed systems failure modes in high-throughput/low-latency environments and fluency in Kubernetes, AWS, and modern observability tooling (Datadog or equivalent).
  • •Ability to read/write production code (Go, TypeScript, or similar) and infrastructure as code; plus a track record of leading through influence without direct reports.
Experience:10+ years
Skills:Incident leadershipLeading through influenceWritten communicationTeaching/coachingPragmatism
Languages:English
Tech Stack:KubernetesAWSDatadogGoTypeScriptElasticsearchRedisDynamoDBKafkaObservabilitySLIsSLOsError budgetsInfra as code

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Fingerprint
Provides passwordless authentication and device intelligence solutions to help businesses verify users and prevent fraud. Offers passkey-based authentication, risk signals, and developer-friendly APIs to replace passwords and improve account security and user experience.
Industry: Cybersecurity
Website