Staff Backend Engineer - Monitoring and Anomaly Detection (Monetization)

GitLab
Bengaluru
Workplace: RemoteFull timeFunction: Software EngineeringSkills: ["Judgment","Reliability focus","Clear communication","Technical direction"]

Lead observability and anomaly detection for the Monetization fulfillment workflow, shaping telemetry, detection, and reconciliation tooling that spots billing, data, and event anomalies before they impact revenue or customer experience. Build and operate metrics/logs/traces using Prometheus, Grafana, and OpenTelemetry, implement automated detection and data integrity checks, define SLOs, and improve reliability through runbooks and incident response. Collaborate with Product, Finance, Support, and engineering teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
GitLab
GitLab
2 days ago

Staff Backend Engineer - Monitoring and Anomaly Detection (Monetization)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Lead observability and anomaly detection for the Monetization fulfillment workflow, shaping telemetry, detection, and reconciliation tooling that spots billing, data, and event anomalies before they impact revenue or customer experience. Build and operate metrics/logs/traces using Prometheus, Grafana, and OpenTelemetry, implement automated detection and data integrity checks, define SLOs, and improve reliability through runbooks and incident response. Collaborate with Product, Finance, Support, and engineering teams.
Location: Bengaluru
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Design, build, and operate observability across the Monetization stack using metrics, logs, and traces with Prometheus and Grafana.
  • •Implement automated detection for billing, data, and event anomalies and route alerts to responsible feature teams.
  • •Develop reconciliation and data integrity checks across usage and billing pipelines.
  • •Define and track service level objectives and indicators; write runbooks and participate in incident response to improve reliability.
  • •Apply AI/ML approaches to predict system anomalies, and review merge requests and feedback with Monetization engineering.

Pay and Benefits

Perks:Paid LeaveEquityLearning BudgetParental Leave

Key Requirements

  • •Professional experience building and operating applications with Ruby on Rails (including CustomersDot).
  • •Experience in site reliability or observability engineering, including monitoring, alerting, SLOs/SI, runbooks, and incident response.
  • •Proven ability to set technical direction for observability, reliability, or detection work and bring other engineers along.
  • •Experience building anomaly detection, monitoring, or risk management tooling.
  • •Familiarity with observability tools such as Prometheus, Grafana, and OpenTelemetry.
Experience:Observability
Skills:JudgmentReliability focusClear communicationTechnical direction
Tech Stack:Ruby on RailsRubyPythonPrometheusGrafanaOpenTelemetryClickHouseSiphonNATS JetStreamMachine LearningSalesforceZuora

Company Brief

GitLab
Provides a single application for the complete DevSecOps lifecycle, offering source code management, CI/CD, security, and collaboration tools to help teams deliver software faster and more securely.
Industry: Developer Tools
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2011
WebsiteLinkedIn