Senior Site Reliability Engineer

Cribl
United States
Workplace: RemoteFull timeUSD 141,800 - 195,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Autonomy","Collaboration","Consensus building","Testing","High quality mindset"]

Design, deploy, and operate reliable observability infrastructure for complex, cloud-based platforms. You’ll work with engineering, product, and platform teams to improve service delivery across the lifecycle, monitor availability/latency/system health, drive incident and stability improvements, and reduce operational toil through automation and innovation. The role requires on-call/off-hours support and uses modern reliability practices, observability tooling, and infrastructure-as-code.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cribl
Cribl
2 months ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Design, deploy, and operate reliable observability infrastructure for complex, cloud-based platforms. You’ll work with engineering, product, and platform teams to improve service delivery across the lifecycle, monitor availability/latency/system health, drive incident and stability improvements, and reduce operational toil through automation and innovation. The role requires on-call/off-hours support and uses modern reliability practices, observability tooling, and infrastructure-as-code.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Engage with teams to improve service delivery and reliability across their entire lifecycle.
  • •Measure and monitor production systems for availability, latency, and overall system health.
  • •Investigate errors and instability in production cloud services and drive operational excellence.
  • •Partner with product and platform teams to evolve systems for reliability, resilience, and observability.
  • •Reduce toil through automation and creative innovation, including standing by/on-call/off-hours duties.

Pay and Benefits

Salary: USD 141,800 - 195,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVisionLife InsurancePaid HolidaysPaid LeaveFertility Treatment401kEquity

Key Requirements

  • •Proven experience designing, implementing, and operating observability systems for complex cloud-based platforms.
  • •Experience with configuration management and infrastructure as code, such as Terraform (preferred) or Ansible.
  • •Knowledge of cloud platforms (prefer AWS and Azure) plus container and orchestration technologies.
  • •Experience with APM/observability tools including New Relic, Splunk, CloudWatch, Prometheus, Grafana/Kibana, and Sentry.
  • •Development experience with JavaScript/Node.js/TypeScript in a Linux/Mac environment, plus incident response in a blameless environment.
Experience:ObservabilityCloud platformsSite reliabilityIncident response
Skills:AutonomyCollaborationConsensus buildingTestingHigh quality mindset
Languages:English
Tech Stack:TerraformAnsibleAWSAzureCloud SDKsContainersOrchestrationNew RelicSplunkCloudWatchPrometheusGrafanaKibanaSentryContinuous deliveryJavaScriptNode.jsTypeScriptLinuxMac

Company Brief

Cribl
Builds observability and data routing software that enables organizations to collect, process, and route machine data (logs, metrics, traces) to analytics and storage destinations while reducing costs and improving operational visibility.
Industry: Data Infrastructure
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Funding: Series D
Headquarters: San Francisco, United States
Founded: 2017
WebsiteLinkedIn