Database Reliability Engineer - Core Team

ClickHouse
EMEA
Workplace: RemoteFull timeFunction: Administration & Executive AssistanceExperience: 5+ yearsSkills: ["Problem-solving","Production debugging","Communication","Ownership","Accountability"]

Build and lead reliability processes for ClickHouse Core, improving reliability, availability, scalability, and performance for ClickHouse Cloud. Own incident response, escalations, investigations, and blameless post-mortems, collaborating with teams across Control Plane, Dataplane, Security, Support, and Operations. Create metrics and alerts, debug production issues affecting customers, and drive chaos initiatives to proactively strengthen resilience across engineering teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ClickHouse
ClickHouse
5 months ago

Database Reliability Engineer - Core Team

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live
Reposted: similar role first listed 1 year ago

Job Summary

Build and lead reliability processes for ClickHouse Core, improving reliability, availability, scalability, and performance for ClickHouse Cloud. Own incident response, escalations, investigations, and blameless post-mortems, collaborating with teams across Control Plane, Dataplane, Security, Support, and Operations. Create metrics and alerts, debug production issues affecting customers, and drive chaos initiatives to proactively strengthen resilience across engineering teams.
Location: EMEA
Workplace: Remote
Employment Type: Full time
Job Function: Administration & Executive Assistance
Seniority: Mid level

Key Responsibilities

  • •Continuously improve the reliability and performance of ClickHouse Core.
  • •Create metrics and alerts to identify and prevent production problems before they affect customers.
  • •Investigate common customer problems, identify root causes, and submit bug fixes and issue reports.
  • •Improve incident response processes and run blameless post-mortems for Core-related outages with support and cloud teams.
  • •Manage on-call processes and drive chaos initiatives across engineering teams to minimize customer impact.

Pay and Benefits

Perks:Health InsuranceEquityHome Office

Key Requirements

  • •Bachelor’s or Master’s degree in Computer Science or a related field.
  • •At least 5 years of experience in Reliability Engineering, QA, or customer-facing engineering.
  • •Previous experience operating ClickHouse or other SQL databases in production.
  • •Excellent understanding of distributed database internals and SQL (ClickHouse is a plus).
  • •Scripting experience with Shell or Python, and ability to read and understand C++ code.
Experience:5+ yearsReliability engineeringQASQL databasesDistributed systems
Education:
Skills:Problem-solvingProduction debuggingCommunicationOwnershipAccountability
Tech Stack:ClickHouseSQLShellPythonC++AWSAzureGoogle Cloud PlatformCloud computing

Company Brief

ClickHouse
Develops ClickHouse, a high-performance open-source columnar database for real-time analytics, enabling fast querying and processing of large volumes of data for analytics, monitoring, and business intelligence workloads.
Industry: Data Infrastructure
Headquarters: Menlo Park, United States
Founded: 2016
WebsiteLinkedIn