Senior Site Reliability Engineer

Gradle
Europe
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsSkills: ["Self-direction","Written communication","Incident management","Automation mindset"]

Help found a new SRE team for a toolchain observability and intelligence platform used by global software organizations. Own reliability for Develocity instances and supporting services, run follow-the-sun incident response, and drive automation across deployment, upgrades, monitoring, and recovery. Build observability (logging, metrics, tracing, alerting), optimize performance and costs, and collaborate across stacks and engineering teams to embed reliability from the start.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Gradle
Gradle
1 month ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Help found a new SRE team for a toolchain observability and intelligence platform used by global software organizations. Own reliability for Develocity instances and supporting services, run follow-the-sun incident response, and drive automation across deployment, upgrades, monitoring, and recovery. Build observability (logging, metrics, tracing, alerting), optimize performance and costs, and collaborate across stacks and engineering teams to embed reliability from the start.
Location: Europe
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Operate and maintain Develocity instances and supporting services.
  • •Participate in follow-the-sun on-call for incident response across the stack.
  • •Drive automation across deployment, upgrades, monitoring, self-healing, and recovery.
  • •Build and maintain observability (logging, metrics, tracing, alerting).
  • •Own disaster recovery, backups, business continuity, and run incident retrospectives.

Pay and Benefits

Perks:EquityRemote Work

Key Requirements

  • •5+ years in SRE, DevOps, or equivalent operating production services at scale.
  • •Strong Kubernetes production experience.
  • •Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).
  • •Proficiency with observability tools (Prometheus, Grafana) and Infrastructure as Code (Terraform).
  • •Track record in incident management and response; also able to do 24/7 on-call rotations.
Experience:5+ yearsSREDevOpsSaaS
Skills:Self-directionWritten communicationIncident managementAutomation mindset
Languages:English
Tech Stack:DevelocityKubernetesAWSEKSRDSS3EC2PrometheusGrafanaTerraformPythonBashLoggingMetricsTracingAlertingInfrastructure as CodeArtifact registries

Company Brief

Gradle
Develops Gradle, a build automation and dependency management tool and Gradle Enterprise for optimizing build and test performance. Serves software engineering teams to accelerate CI/CD, improve developer productivity, and scale build infrastructure.
Industry: Developer Tools
Company Size: Medium (51 to 250 employees)
Growth: Growth Stage Startup
Headquarters: San Francisco, United States
Founded: 2007
WebsiteLinkedIn