Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)

Crowdstrke
Redmond, Sunnyvale, Austin
Workplace: HybridFull timeUSD 140,000 - 215,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsSkills: ["Systems thinking","Team collaboration","Mentorship","Technical leadership","Ownership"]

Build and evolve core reliability libraries, services, and tooling used across the Falcon platform. Partner with product engineering leadership to deliver multi-year reliability roadmaps—improving scalability, performance, observability, and automation. Write production code, re-architect critical systems, and lead resilience engineering (chaos/failure injection) while defining SLOs and error budgets. Help eliminate manual toil with infrastructure-as-code and shared platform components.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Crowdstrke
Crowdstrke
3 days ago

Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 45 minutes agoStatus: Live

Job Summary

Build and evolve core reliability libraries, services, and tooling used across the Falcon platform. Partner with product engineering leadership to deliver multi-year reliability roadmaps—improving scalability, performance, observability, and automation. Write production code, re-architect critical systems, and lead resilience engineering (chaos/failure injection) while defining SLOs and error budgets. Help eliminate manual toil with infrastructure-as-code and shared platform components.
Location: Redmond, Sunnyvale, Austin
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Partner with engineering leadership to define and drive multi-year reliability roadmaps across product groups.
  • •Design and implement architectural improvements to services, libraries, and platforms used organization-wide.
  • •Develop and maintain services meeting aggressive reliability, scalability, and performance demands.
  • •Build cross-cutting libraries for the cloud platform and establish foundational observability practices (tracing, profiling, alerting, SLOs/SLIs) and error budgets.
  • •Lead reliability/scalability/performance/cost-efficiency initiatives, including resilience engineering and infrastructure-as-code automation; provide technical leadership during complex incidents and ensure follow-through.

Pay and Benefits

Salary: USD 140,000 - 215,000 annually
Equity and Bonus:Equity
Perks:Health Insurance401kPaid LeaveParental LeaveLearning Budget

Key Requirements

  • •10+ years building and operating distributed systems and service-oriented backends at scale.
  • •5+ years developing microservices for a SaaS product in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js).
  • •Expert-level proficiency in at least one programming language, with expert-level Go or demonstrated ability/willingness to reach expert level in Go.
  • •Deep understanding of distributed systems (consensus, replication, consistency models, failure modes, scalability patterns) and scaling backend systems (sharding, partitioning, horizontal scaling, capacity planning, performance optimization).
  • •Track record making impactful architectural decisions at organizational scope and delivering to production, with strong systems thinking and resilient engineering best practices.
Experience:10+ yearsDistributed systemsMicroservicesSaaSCloudCybersecurity
Skills:Systems thinkingTeam collaborationMentorshipTechnical leadershipOwnership
Tech Stack:GoJavaScalaKotlinPythonNode.jsKafkaProtobufUnified SearchAWSCassandraKubernetesElasticsearchOpenSearchGoogle Cloud PlatformGCPOracle Cloud InfrastructureOCI

Company Brief

Crowdstrke
Provides cloud-native endpoint protection, threat intelligence, and security operations solutions that prevent breaches and stop sophisticated cyberattacks across endpoints, cloud workloads, identity, and APIs for enterprises worldwide.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Sunnyvale, United States
Founded: 2011
Glassdoor
Glassdoor: 4.4
WebsiteLinkedIn