Senior Site Reliability Engineer, Observability

Ripple
Chicago, New York
Workplace: OnsiteFull timeUSD 160,000 - 200,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 7+ yearsSkills: ["Coaching","Mentoring","Troubleshooting","Collaboration","Stakeholder communication"]

Design and implement observability and reliability capabilities for enterprise treasury infrastructure, partnering with teams to improve operational maturity. Build monitoring, dashboards, SLOs/SLIs, and alert configurations in New Relic across Azure and AWS, while authoring Terraform-based instrumentation and governance. Help establish an early-stage incident management program using Incident.IO, track MTTR/MTTD, and coach stream-aligned product teams on incident workflows and dashboard/alert best practices.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Ripple
Ripple
1 month ago

Senior Site Reliability Engineer, Observability

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Design and implement observability and reliability capabilities for enterprise treasury infrastructure, partnering with teams to improve operational maturity. Build monitoring, dashboards, SLOs/SLIs, and alert configurations in New Relic across Azure and AWS, while authoring Terraform-based instrumentation and governance. Help establish an early-stage incident management program using Incident.IO, track MTTR/MTTD, and coach stream-aligned product teams on incident workflows and dashboard/alert best practices.
Location: Chicago, New York
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and implement monitoring, alerting, and dashboards in New Relic across Azure and AWS, including NRQL queries for troubleshooting and analysis.
  • •Define and implement SLOs/SLIs and error budgets, and lead alert noise reduction to ensure alerts are actionable.
  • •Develop and maintain Terraform infrastructure for provisioning monitoring resources and observability infrastructure, and author Terraform-based configurations.
  • •Administer Incident.IO for alert routing, notification workflows, runbook management, and establish incident management foundations (PIRs, on-call rotations, escalation policies, response playbooks).
  • •Partner with engineering teams to improve observability maturity (structured logging, metrics instrumentation, distributed tracing, dashboard patterns) and coach teams via workshops and knowledge sharing.

Pay and Benefits

Salary: USD 160,000 - 200,000 annually
Equity and Bonus:Equity
Perks:Mobile PhoneR&r DaysWellness StipendPaid LeaveParental LeaveMeal Allowance

Key Requirements

  • •7+ years in Site Reliability Engineering, DevOps, or Platform Engineering focused on observability and production operations.
  • •Expert-level hands-on experience with New Relic (APM, Infrastructure, Logs, Synthetics, Alerts) and strong NRQL proficiency.
  • •Experience defining and implementing SLOs/SLIs and error budgets for reliability management.
  • •Strong Terraform experience for infrastructure as code and familiarity with IaC governance patterns.
  • •Proficiency with PowerShell scripting, plus experience with Azure cloud and Azure DevOps CI/CD pipelines.
Experience:7+ yearsObservabilityProduction operationsIncident managementCloud infrastructure
Skills:CoachingMentoringTroubleshootingCollaborationStakeholder communication
Certifications:SOC 2ISO 27001
Languages:English
Tech Stack:New RelicNRQLTerraformSLOSLIIncident.IOSlackOpsGeniePagerDutyAzureAWSApp ServicesVirtual MachinesAzure SQLAzure DevOpsOctopus DeployPowerShellWindowsLinuxAzure DevOps pipelines

Company Brief

Ripple
Builds blockchain-based payment and liquidity infrastructure for financial institutions and enterprises. Its products support cross-border payments, crypto liquidity, and digital asset settlement.
Industry: Fintech Infrastructure
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn