Sr. Engineer, DevOps, NG-SIEM (Remote, AUS)

Crowdstrke
Sydney
Workplace: RemoteFull timeFunction: Hospitality & Food ServiceSkills: ["Reliability engineering","Operational excellence","Incident response","Communication","Collaboration"]

Build and operate the NG-SIEM serverless cell platform for Falcon, ensuring health, performance, and reliability at massive scale. Monitor and respond to incidents via follow-the-sun on-call, run capacity planning and scaling operations, and perform cluster upgrades using ring-based rollouts. Develop Go microservices and automation to enable self-healing, while creating SLIs/SLOs, alerting, synthetic tests, and operational dashboards for fast troubleshooting and root-cause analysis.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Crowdstrke
Crowdstrke
1 month ago

Sr. Engineer, DevOps, NG-SIEM (Remote, AUS)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 25 days agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Build and operate the NG-SIEM serverless cell platform for Falcon, ensuring health, performance, and reliability at massive scale. Monitor and respond to incidents via follow-the-sun on-call, run capacity planning and scaling operations, and perform cluster upgrades using ring-based rollouts. Develop Go microservices and automation to enable self-healing, while creating SLIs/SLOs, alerting, synthetic tests, and operational dashboards for fast troubleshooting and root-cause analysis.
Location: Sydney
Workplace: Remote
Employment Type: Full time
Job Function: Hospitality & Food Service
Seniority: Mid level

Key Responsibilities

  • •Monitor and maintain LogScale cell health, performance, and reliability; respond to incidents and participate in follow-the-sun on-call with escalation paths.
  • •Upgrade LogScale clusters and microservices, deploying patches, updating configurations, and recovering cells when failures occur.
  • •Build monitoring for SLIs/SLOs, including alerting, synthetic test suites, and operational dashboards for troubleshooting and root-cause analysis.
  • •Write Go microservices and automation scripts that reduce toil and enable self-healing workflows.
  • •Coordinate with product engineering and customer support to troubleshoot production issues and implement coordinated platform changes.

Pay and Benefits

Equity and Bonus:Equity
Perks:Wellness StipendEquity AwardsParental Leave

Key Requirements

  • •7+ years in site reliability engineering, platform engineering, or infrastructure operations, improving distributed systems at scale.
  • •Experience with capacity planning, resource optimization, and scaling operations for large-scale infrastructure.
  • •Proficiency in Go (or Python/Java/similar) for writing automation scripts and microservices.
  • •Experience with cloud infrastructure (AWS, OCI, or GCP) and infrastructure-as-code (Terraform, Pulumi, or similar).
  • •Strong communication for runbooks and incident reports, plus the ability to balance operational stability with feature delivery.
Experience:Site reliability engineeringPlatform engineeringInfrastructure operationsDistributed systemsCloud infrastructure
Skills:Reliability engineeringOperational excellenceIncident responseCommunicationCollaboration
Tech Stack:GoPythonJavaAWSOCIGCPTerraformPulumiLogScaleMicroservicesInfrastructure as codeRing-based rollout

Company Brief

Crowdstrke
Provides cloud-native endpoint protection, threat intelligence, and security operations solutions that prevent breaches and stop sophisticated cyberattacks across endpoints, cloud workloads, identity, and APIs for enterprises worldwide.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Sunnyvale, United States
Founded: 2011
Glassdoor
Glassdoor: 4.4
WebsiteLinkedIn