Site Reliability Engineer (SRE/ DevOps) - Engineering Productivity

Arista Networks
Dublin
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsEducation: bachelorsSkills: ["Problem-solving","Software troubleshooting","Incident response triage","Communication"]

Build, deploy, and operate secure production systems that are scalable, reliable, observable, performant, and secure. Automate workflows to reduce toil, improve alerting and incident response, and maintain runbooks and postmortems to prevent repeat failures. Partner with product development teams to remove infrastructure bottlenecks and enhance developer experience in a hybrid cloud environment using modern automation, CI/CD, and monitoring tools.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Arista Networks
Arista Networks
10 hours ago

Site Reliability Engineer (SRE/ DevOps) - Engineering Productivity

âś“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Build, deploy, and operate secure production systems that are scalable, reliable, observable, performant, and secure. Automate workflows to reduce toil, improve alerting and incident response, and maintain runbooks and postmortems to prevent repeat failures. Partner with product development teams to remove infrastructure bottlenecks and enhance developer experience in a hybrid cloud environment using modern automation, CI/CD, and monitoring tools.
Location: Dublin
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Build, deploy safely and incrementally, and operate critical production systems focused on scalability, reliability, observability, performance, and security.
  • •Automate workflows to remove toil and enhance alerts with automated handling.
  • •Create and maintain incident response runbooks, triage platform/infrastructure issues, and write postmortems to prevent recurring incidents.
  • •Plan and communicate maintenance windows on production systems and coordinate with third-party vendor support as needed.
  • •Partner with product development teams to identify infrastructure bottlenecks and design solutions to improve developer experience and workflow efficiency.

Key Requirements

  • •BSc Computer Science or Engineering + 3 years’ experience, or MS Computer Science or Engineering + 3 years’ experience, or equivalent work experience.
  • •Knowledge of one or more of Go, Python, or shell scripting to implement medium complexity automation workflows.
  • •Knowledge of Linux (or UNIX) from an administration and debugging perspective.
  • •Hands-on experience operating software systems (infrastructure or complex applications) at scale.
  • •Experience with infrastructure-as-code and server provisioning (including storage/networking perspective).
Experience:3+ years
Education:Bachelor's in Computer Science or Engineering
Skills:Problem-solvingSoftware troubleshootingIncident response triageCommunication
Languages:English (UK)
Tech Stack:GoPythonShell scriptingLinuxAnsibleDockerKubernetesGrafanaSpinnakerMySQLElasticSearchGoogle CloudVarnishPerforceGerritJenkinsArtifactoryPrometheusLokiTempo

Company Brief

Arista Networks
Designs and sells high-performance cloud networking hardware and software for large-scale data centers and enterprise environments, including switches, routers, and network operating systems focused on programmability, telemetry, and low-latency networking.
Industry: Networking Equipment
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 2004
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn