Systems Engineer

Cloudflare
Austin
Workplace: HybridFull timeFunction: IT Operations (Systems/Network Admin)Skills: ["Communication","Collaboration"]

Build and maintain tooling for Cloudflare’s SLO team, enabling engineers to measure service and feature reliability at massive scale. You’ll develop distributed systems, create and run production testing infrastructure, and help define SLO/SLI-based monitoring and availability reporting. Partner with product managers and Product Site Reliability Engineers to improve quality-of-service measurements for enterprise customers, collaborating across teams to prevent incidents and accelerate safe delivery.

This position is no longer accepting applications.

  • See live roles at Cloudflare
  • Search all live jobs
  • Browse companies, collections, and locations hiring now
Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa

This position is no longer accepting applications.

See live roles at CloudflareSearch all live jobsBrowse companies, collections, and locations hiring now

Cloudflare
Cloudflare
1 month ago

Systems Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Closed
Reposted: similar role first listed 7 months ago

Job Summary

Build and maintain tooling for Cloudflare’s SLO team, enabling engineers to measure service and feature reliability at massive scale. You’ll develop distributed systems, create and run production testing infrastructure, and help define SLO/SLI-based monitoring and availability reporting. Partner with product managers and Product Site Reliability Engineers to improve quality-of-service measurements for enterprise customers, collaborating across teams to prevent incidents and accelerate safe delivery.
Location: Austin
Workplace: Hybrid
Employment Type: Full time
Job Function: IT Operations (Systems/Network Admin)

Key Responsibilities

  • •Build and run tooling and an internal platform that helps other engineering teams measure service and feature reliability.
  • •Develop and maintain secure, highly available distributed systems in support of production reliability needs.
  • •Create and maintain production testing infrastructure and availability reporting.
  • •Develop and execute test and SLO plans, test cases, and test scripts to verify systems continue operating as expected.
  • •Collaborate with engineers, product managers, and Product Site Reliability Engineers to deliver quality-of-service measurements for enterprise customers.

Key Requirements

  • •Proven track record as a software engineer or similar role.
  • •Programming experience with Go, Rust, or Python.
  • •Experience designing, implementing, and maintaining secure, highly available distributed systems.
  • •Ability to develop, document, and execute test and SLO plans, including test cases and scripts.
  • •Experience measuring uptime metrics such as correctness, availability, and latency SLOs/SLIs.
Experience:Distributed systemsSLOs/SLIsProduction testingReliability engineering
Skills:CommunicationCollaboration
Languages:English
Tech Stack:GoRustPythonClickhousePrometheusGraphQLPostgresSLOs/SLIsLoad testing

Company Brief

Cloudflare
Provides a global network and cloud platform that delivers security, performance, and reliability services for web applications, APIs, and Internet properties, including CDN, DDoS protection, DNS, and zero-trust security solutions.
Industry: Cybersecurity
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 2009
WebsiteLinkedIn