Cloud Software Engineer - Observability Platform

ClickHouse
United States
Workplace: RemoteFull timeUSD 141,000 - 208,000 annuallyFunction: Software EngineeringExperience: 5+ yearsSkills: ["Ownership","Debugging","Clear communication","Pragmatic decision-making","Incident response"]

Build and operate distributed telemetry systems at ClickHouse’s Observability Platform and Internal Observability teams. You’ll design high-scale ingestion, processing, buffering, storage, and autoscaling, own reliability/performance/cost for telemetry pipelines, and participate in on-call to resolve incidents and drive root-cause fixes. You’ll automate repetitive operational work, collaborate with product/infrastructure/service teams, and shape the roadmap through bottleneck analysis and architecture reviews.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
ClickHouse
ClickHouse
5 days ago

Cloud Software Engineer - Observability Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live
Reposted: similar role first listed 10 months ago

Job Summary

Build and operate distributed telemetry systems at ClickHouse’s Observability Platform and Internal Observability teams. You’ll design high-scale ingestion, processing, buffering, storage, and autoscaling, own reliability/performance/cost for telemetry pipelines, and participate in on-call to resolve incidents and drive root-cause fixes. You’ll automate repetitive operational work, collaborate with product/infrastructure/service teams, and shape the roadmap through bottleneck analysis and architecture reviews.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Design, build, and operate distributed systems that ingest, process, and store telemetry at very high scale.
  • •Own reliability, performance, capacity, and cost-efficiency of telemetry pipelines and storage systems.
  • •Participate in on-call rotation to resolve production incidents and drive root-cause fixes to completion.
  • •Build software and automation to eliminate repetitive operational work and improve platform operability.
  • •Identify architectural bottlenecks and help shape the roadmap for the next stage of scale.

Pay and Benefits

Salary: USD 141,000 - 208,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceEquityHome OfficeFlexible Time

Key Requirements

  • •5+ years building and operating production systems at scale.
  • •Strong proficiency in Go.
  • •Experience building and operating services on Kubernetes.
  • •Hands-on infrastructure-as-code and GitOps tooling such as Terraform, Helm, and Argo CD.
  • •Production experience with a major cloud provider (AWS, GCP, or Azure) and telemetry systems (OpenTelemetry, Prometheus, Grafana, or similar).
Experience:5+ years
Skills:OwnershipDebuggingClear communicationPragmatic decision-makingIncident response
Languages:English
Tech Stack:GoKubernetesTerraformHelmArgo CDAWSGCPAzureOpenTelemetryPrometheusGrafanaTypeScript

Company Brief

ClickHouse
Develops ClickHouse, a high-performance open-source columnar database for real-time analytics, enabling fast querying and processing of large volumes of data for analytics, monitoring, and business intelligence workloads.
Industry: Data Infrastructure
Headquarters: Menlo Park, United States
Founded: 2016
WebsiteLinkedIn