Site Reliability Engineer

SingleStore
Lisbon
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Cross-group collaboration","Communication","Problem-solving","Debugging","Troubleshooting"]

Build automation and operational tooling to manage infrastructure rollouts across major cloud providers. Improve telemetry and monitoring to detect customer-impacting events, drive debugging, and enhance the overall customer experience. Partner with engineering teams to optimize service performance and diagnose live site issues, including postmortems and RCA. Join an SLA-driven on-call rotation with after-hours and weekend coverage in a globally distributed environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
SingleStore
SingleStore
3 days ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 minute agoStatus: Live

Job Summary

Build automation and operational tooling to manage infrastructure rollouts across major cloud providers. Improve telemetry and monitoring to detect customer-impacting events, drive debugging, and enhance the overall customer experience. Partner with engineering teams to optimize service performance and diagnose live site issues, including postmortems and RCA. Join an SLA-driven on-call rotation with after-hours and weekend coverage in a globally distributed environment.
Location: Lisbon
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Develop an automation platform to manage infrastructure rollouts across cloud providers.
  • •Optimize telemetry to identify customer-impacting events and provide relevant data for debugging.
  • •Partner with engineering teams to optimize service performance for cloud architectures.
  • •Debug live site events and conduct follow-up postmortems with RCA analysis.
  • •Participate in an SLA-driven on-call rotation including after-hours, weekends, and rotating holidays.

Key Requirements

  • •Infrastructure automation experience; Python and Golang are a plus.
  • •Knowledge of Kubernetes and the container ecosystem.
  • •Experience debugging, diagnosing, and troubleshooting complex production software.
  • •Familiarity with at least one of AWS, Azure, or Google Cloud.
  • •A B.S. degree in Computer Science or related field.
Education:Bachelor's in Computer Science
Skills:Cross-group collaborationCommunicationProblem-solvingDebuggingTroubleshooting
Languages:English
Tech Stack:PythonGolangKubernetesAWSAzureGoogle CloudCloud providersMonitoring

Company Brief

SingleStore
Provides a distributed, high-performance SQL database for real-time analytics and transactions. SingleStore delivers a cloud-native platform that combines transactional and analytical workloads for modern data-intensive applications.
Industry: Data Infrastructure
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2011
WebsiteLinkedIn