Senior Site Reliability Engineer - CloudVision

Arista Networks
Dublin
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Analytical rigor","Problem-solving","Methodical troubleshooting","Cross-functional collaboration","Clear technical communication"]

Design, build, and operate scalable, reliable, observable, and secure production systems. Automate infrastructure to reduce toil, monitor systems with intelligent alerting, and run incident response with runbooks and postmortems. Partner with software engineering to remove infrastructure bottlenecks, manage monitoring stacks, and execute staged deployments and maintenance windows. Work from Ireland on a permanent remote basis, leveraging tools like Kubernetes, Docker, Prometheus, and Grafana.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Arista Networks
Arista Networks
2 days ago

Senior Site Reliability Engineer - CloudVision

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live
Reposted: similar role first listed 2 weeks ago

Job Summary

Design, build, and operate scalable, reliable, observable, and secure production systems. Automate infrastructure to reduce toil, monitor systems with intelligent alerting, and run incident response with runbooks and postmortems. Partner with software engineering to remove infrastructure bottlenecks, manage monitoring stacks, and execute staged deployments and maintenance windows. Work from Ireland on a permanent remote basis, leveraging tools like Kubernetes, Docker, Prometheus, and Grafana.
Location: Dublin
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design, build, and deploy production systems focused on scalability, reliability, observability, performance, and security.
  • •Develop and maintain automation solutions to eliminate toil and streamline operations across production environments.
  • •Monitor production systems, define alerting strategies, and implement automated incident response to minimize downtime.
  • •Create and maintain incident response runbooks and perform postmortems to identify root causes and prevent recurrence.
  • •Collaborate with software engineering teams to resolve infrastructure bottlenecks and improve product deployment workflows.

Key Requirements

  • •Bachelor's degree in Computer Science, Engineering, or equivalent professional experience (5+ years in an infrastructure or systems role).
  • •Proficiency in one or more languages: Go, Python, or bash shell scripting to implement medium-complexity automation workflows.
  • •Strong Linux/UNIX knowledge for administration and debugging, with hands-on experience operating systems at scale in production.
  • •Expertise in infrastructure-as-code principles and practices.
  • •Experience with incident response, postmortem analysis, and continuous improvement methodologies.
Experience:5+ yearsInfrastructureSystemsProduction operationsIncident responseObservability
Education:Bachelor's
Skills:Analytical rigorProblem-solvingMethodical troubleshootingCross-functional collaborationClear technical communication
Languages:English (UK)
Tech Stack:AnsibleTerraformGoPythonBashLinuxUNIXPrometheusGrafanaAWSGCPAzureKubernetesDockerCI/CDJenkinsGitLabOpen-sourcePostgreSQLSpinnaker

Company Brief

Arista Networks
Designs and sells high-performance cloud networking hardware and software for large-scale data centers and enterprise environments, including switches, routers, and network operating systems focused on programmability, telemetry, and low-latency networking.
Industry: Networking Equipment
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 2004
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn