Senior Site Reliability Engineer- Remote

ClickHouse
United States
Workplace: RemoteFull timeUSD 141,000 - 208,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 8+ yearsEducation: bachelorsSkills: ["Communication","Interpersonal skills","Problem-solving","Ownership","Team collaboration"]

Join ClickHouse as a Senior Site Reliability Engineer to design, build, and operate scalable, highly available cloud infrastructure for ClickHouse Cloud. You will own incident management, SLO/SLAs, on-call processes, and drive reliability and performance improvements across Dataplane, Control Plane, and Core components, collaborating with multiple engineering teams in a fast-paced, remote-first environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ClickHouse
ClickHouse
4 months ago

Senior Site Reliability Engineer- Remote

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Join ClickHouse as a Senior Site Reliability Engineer to design, build, and operate scalable, highly available cloud infrastructure for ClickHouse Cloud. You will own incident management, SLO/SLAs, on-call processes, and drive reliability and performance improvements across Dataplane, Control Plane, and Core components, collaborating with multiple engineering teams in a fast-paced, remote-first environment.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Sr. Manager level

Key Responsibilities

  • •Collaborate with various engineering teams in ClickHouse to design and implement scalable, secure, and highly available systems for ClickHouse.
  • •Establish and manage service level objectives (SLOs) and service level agreements (SLAs) for ClickHouse Cloud.
  • •Ensure all the infrastructure components in ClickHouse Cloud have monitoring and alerting in place to ensure timely detection and resolution of incidents.
  • •Enhance and refine incident response processes and post-mortem analysis for outages in ClickHouse Cloud including communicating to customers.
  • •Continuously improve the reliability and performance of ClickHouse services.

Pay and Benefits

Salary: USD 141,000 - 208,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceEquityHome OfficeTime OffRemote Work

Key Requirements

  • •Bachelor’s or Master’s degree in Computer Science or a related field
  • •At least 8 years of experience in Site Reliability Engineering or a related field
  • •Previous experience using ClickHouse in production
  • •Hands on experience with Go and/or Python
  • •Strong knowledge of cloud computing platforms such as AWS, Azure, or Google Cloud Platform
Experience:8+ yearsCloud computingDistributed systemsServerlessAnalyticsCloud infrastructure
Education:Bachelor's
Skills:CommunicationInterpersonal skillsProblem-solvingOwnershipTeam collaboration
Languages:English
Tech Stack:GoPythonAWSAzureGoogle Cloud PlatformKubernetesDockerAnsibleTerraformPuppetClickHouseSQL

Company Brief

ClickHouse
Develops ClickHouse, a high-performance open-source columnar database for real-time analytics, enabling fast querying and processing of large volumes of data for analytics, monitoring, and business intelligence workloads.
Industry: Data Infrastructure
Headquarters: Menlo Park, United States
Founded: 2016
WebsiteLinkedIn