Engineer II, Site Reliability (Remote, GBR)

Crowdstrke
United Kingdom
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Ownership","Urgency","Analytical skills","Communication","Incident response"]

Design, operate, and improve CrowdStrike’s Commercial Cloud reliability as an Engineer II on the TechOps SRE team. You’ll ensure platform availability, latency, throughput, monitoring, incident response, and capacity planning for large-scale distributed systems. Work with Linux at scale, modern telemetry stacks (ELK, Prometheus, Grafana, Zabbix), configuration management tools (Puppet/Chef/Ansible), and storage/infrastructure technologies to drive 24x7 excellence.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Crowdstrke
Crowdstrke
4 days ago

Engineer II, Site Reliability (Remote, GBR)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Design, operate, and improve CrowdStrike’s Commercial Cloud reliability as an Engineer II on the TechOps SRE team. You’ll ensure platform availability, latency, throughput, monitoring, incident response, and capacity planning for large-scale distributed systems. Work with Linux at scale, modern telemetry stacks (ELK, Prometheus, Grafana, Zabbix), configuration management tools (Puppet/Chef/Ansible), and storage/infrastructure technologies to drive 24x7 excellence.
Location: United Kingdom
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Operate the platform for availability, latency, throughput, monitoring, issue response, and capacity planning.
  • •Troubleshoot server hardware issues and ensure the platform operates flawlessly 24x7.
  • •Lead incident analysis and drive incident response practices toward resolution.
  • •Develop automation and tooling through software to support mission-critical services for large-scale distributed systems.
  • •Gather and analyze metrics from operating systems and applications to support performance tuning and fault finding.

Pay and Benefits

Equity and Bonus:Equity
Perks:EquityWellness StipendPaid LeaveParental Leave

Key Requirements

  • •Bachelor's degree and/or equivalent experience in Computer Science.
  • •Minimum five years’ experience in a large-scale production environment.
  • •Minimum two years’ experience in software engineering and in one or more of: C++, Java, Python, Go.
  • •Experience with infrastructure technologies such as Linux, Windows, VMware, Docker, and Kubernetes (or similar).
  • •Experience with monitoring/telemetry and incident response, plus configuration management using tools like Puppet, Chef, or Ansible.
Experience:5+ yearsDistributed systemsProduction environments
Education:Bachelor's
Skills:OwnershipUrgencyAnalytical skillsCommunicationIncident response
Tech Stack:LinuxC++JavaPythonGoSANNASNFSObject StorageFreeNASISCSIWindowsVMwareDockerKubernetesELKPrometheusGrafanaZabbixPuppet

Company Brief

Crowdstrke
Provides cloud-native endpoint protection, threat intelligence, and security operations solutions that prevent breaches and stop sophisticated cyberattacks across endpoints, cloud workloads, identity, and APIs for enterprises worldwide.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Sunnyvale, United States
Founded: 2011
Glassdoor
Glassdoor: 4.4
WebsiteLinkedIn