Site Reliability Engineer, Observability

Ripple
Chicago, New York
Workplace: HybridFull timeUSD 160,000 - 200,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Coaching","Mentoring","Troubleshooting","Collaboration","Communication"]

Build and run observability and reliability systems for enterprise treasury payments across Azure and AWS. Design New Relic monitoring, alerting, dashboards, SLOs/SLIs, and error budgets; reduce alert noise and optimize observability costs. Develop observability infrastructure with Terraform and author/troubleshoot Azure DevOps pipelines. Establish and expand incident management foundations with Incident.IO, integrating Slack and OpsGenie, and coach stream-aligned product teams to improve operational maturity.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Ripple
Ripple
3 months ago

Site Reliability Engineer, Observability

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live
Reposted: similar role first listed 3 months ago

Job Summary

Build and run observability and reliability systems for enterprise treasury payments across Azure and AWS. Design New Relic monitoring, alerting, dashboards, SLOs/SLIs, and error budgets; reduce alert noise and optimize observability costs. Develop observability infrastructure with Terraform and author/troubleshoot Azure DevOps pipelines. Establish and expand incident management foundations with Incident.IO, integrating Slack and OpsGenie, and coach stream-aligned product teams to improve operational maturity.
Location: Chicago, New York
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and implement New Relic monitoring, alerting, and dashboards across Azure and AWS, including NRQL queries for troubleshooting, analysis, and reporting.
  • •Define and implement SLOs/SLIs and error budgets, and lead alert noise reduction to ensure alerts are actionable.
  • •Develop and maintain Terraform-based observability infrastructure and establish IaC governance standards across teams.
  • •Administer Incident.IO for alert routing, notification workflows, runbook management, and build incident management foundations (on-call, escalation, severity, playbooks).
  • •Coach and consult with stream-aligned product teams through workshops and hands-on guidance to improve observability maturity and incident response practices.

Pay and Benefits

Salary: USD 160,000 - 200,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceRetirementLearning BudgetMobile PhoneWellness StipendParental Leave

Key Requirements

  • •7+ years in Site Reliability Engineering, DevOps, or Platform Engineering focused on observability and production operations.
  • •Expert-level hands-on experience with New Relic (APM, Infrastructure, Logs, Synthetics, Alerts) and strong NRQL proficiency for troubleshooting and analysis.
  • •Experience defining and implementing SLOs/SLIs and error budgets, plus incident response workflows such as on-call rotations, escalation policies, and post-incident reviews.
  • •Strong Terraform experience for IaC covering monitoring resources and governance patterns.
  • •Proficiency with PowerShell scripting and strong experience with Azure cloud (App Services, Virtual Machines, Azure SQL, networking, monitoring), with working knowledge of AWS.
Experience:ObservabilityIncident managementSREFinTech
Skills:CoachingMentoringTroubleshootingCollaborationCommunication
Languages:English
Tech Stack:New RelicNRQLTerraformAzureAWSIncident.IOSlackOpsGeniePagerDutyAzure DevOpsOctopus DeployPowerShellSLOsSLIsDistributed tracingStructured logging

Company Brief

Ripple
Builds blockchain-based payment and liquidity infrastructure for financial institutions and enterprises. Its products support cross-border payments, crypto liquidity, and digital asset settlement.
Industry: Fintech Infrastructure
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn