Senior Site Reliability Engineer, Observability

Ripple
New York, Chicago
Workplace: OnsiteFull timeUSD 160,000 - 200,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 7+ yearsSkills: ["Coaching","Mentoring","Troubleshooting","Stakeholder communication","Cross-functional collaboration"]

Own observability and reliability engineering for enterprise treasury infrastructure, designing monitoring, alerting, dashboards, and New Relic NRQL queries. Define SLOs/SLIs and error budgets, reduce alert noise, and optimize observability costs. Build and govern observability infrastructure using Terraform and incident management foundations with Incident.IO, Slack, and OpsGenie integrations. Coach stream-aligned product teams and improve incident response outcomes across Azure and AWS environments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Ripple
Ripple
1 month ago

Senior Site Reliability Engineer, Observability

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 hours agoStatus: Live

Job Summary

Own observability and reliability engineering for enterprise treasury infrastructure, designing monitoring, alerting, dashboards, and New Relic NRQL queries. Define SLOs/SLIs and error budgets, reduce alert noise, and optimize observability costs. Build and govern observability infrastructure using Terraform and incident management foundations with Incident.IO, Slack, and OpsGenie integrations. Coach stream-aligned product teams and improve incident response outcomes across Azure and AWS environments.
Location: New York, Chicago
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design and implement monitoring, alerting, and dashboards in New Relic across Azure and AWS, including NRQL query authoring for troubleshooting and reporting.
  • •Define and implement SLOs/SLIs and error budgets; coach teams on using them to balance velocity with reliability.
  • •Lead alert noise reduction and signal quality engineering by tuning thresholds and eliminating false positives.
  • •Develop and maintain Terraform-based IaC for provisioning and managing observability monitoring and alerting resources.
  • •Administer incident management tooling (Incident.IO), build incident response foundations (postmortems, on-call rotation, escalation policies), and support production incidents with debriefs and continuous improvement.

Pay and Benefits

Salary: USD 160,000 - 200,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceRetirementLearning BudgetMobile PhoneParental LeaveWellness Stipend

Key Requirements

  • •7+ years in Site Reliability Engineering, DevOps, or Platform Engineering with a strong focus on observability and production operations.
  • •Expert hands-on New Relic experience (APM, Infrastructure, Logs, Synthetics, Alerts) and strong NRQL proficiency for troubleshooting and analysis.
  • •Expertise defining and implementing SLOs/SLIs and error budgets, plus experience designing incident response workflows and on-call rotations.
  • •Strong Terraform experience for infrastructure-as-code for cloud and monitoring resources, including IaC governance patterns.
  • •Proficiency with PowerShell and strong Azure experience (App Services, Virtual Machines, Azure SQL, networking, monitoring), with working knowledge of AWS.
Experience:7+ yearsFinTech
Skills:CoachingMentoringTroubleshootingStakeholder communicationCross-functional collaboration
Certifications:SOC 2ISO 27001
Languages:English
Tech Stack:New RelicNRQLTerraformIncident.IOSlackOpsGenieAzureAWSWindowsLinuxPowerShellAzure DevOpsOctopus DeployApp ServicesVirtual MachinesAzure SQLDistributed tracingSLOSLIMTTR

Company Brief

Ripple
Builds blockchain-based payment and liquidity infrastructure for financial institutions and enterprises. Its products support cross-border payments, crypto liquidity, and digital asset settlement.
Industry: Fintech Infrastructure
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Headquarters: San Francisco, United States
Founded: 2012
WebsiteLinkedIn