Senior Site Reliability Engineer - CloudVision

Arista Networks
Dublin
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 5+ yearsEducation: bachelorsSkills: ["Analytical thinking","Problem-solving","Methodical troubleshooting","Collaboration","Clear communication"]

Build, deploy, and operate scalable production systems with strong reliability, observability, and security. Automate operations to reduce toil, monitor services with intelligent alerting, and drive incident response using runbooks and postmortems. Partner with software teams to remove infrastructure bottlenecks, optimize monitoring stacks, and plan low-disruption maintenance. Work hands-on with infrastructure-as-code and cloud platforms across AWS, GCP, and Azure.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Arista Networks
Arista Networks
3 days ago

Senior Site Reliability Engineer - CloudVision

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Build, deploy, and operate scalable production systems with strong reliability, observability, and security. Automate operations to reduce toil, monitor services with intelligent alerting, and drive incident response using runbooks and postmortems. Partner with software teams to remove infrastructure bottlenecks, optimize monitoring stacks, and plan low-disruption maintenance. Work hands-on with infrastructure-as-code and cloud platforms across AWS, GCP, and Azure.
Location: Dublin
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Design, build, and deploy production systems with scalability, reliability, observability, and performance while meeting security standards.
  • •Develop and maintain automation to eliminate toil and improve operational efficiency across production environments.
  • •Proactively monitor systems, create alerting strategies, and implement automated incident response to minimize downtime.
  • •Create and maintain incident response runbooks and conduct postmortems to identify root causes and prevent recurrence.
  • •Collaborate with software engineering teams to resolve infrastructural bottlenecks and improve product deployment workflows, including staged safe rollouts.

Key Requirements

  • •Bachelor's degree in Computer Science, Engineering, or equivalent experience with 5+ years in infrastructure or systems roles.
  • •Programming proficiency in Go, Python, or bash shell scripting for medium-complexity automation workflows.
  • •Strong Linux/UNIX knowledge for administration and debugging.
  • •Hands-on experience operating software systems and infrastructure at scale in production environments.
  • •Expertise in infrastructure-as-code, incident response, and continuous improvement methodologies.
Experience:5+ yearsInfrastructureSystemsProduction operationsObservabilityIncident response
Education:Bachelor's
Skills:Analytical thinkingProblem-solvingMethodical troubleshootingCollaborationClear communication
Languages:English (UK)
Tech Stack:GoPythonBashLinuxUNIXInfrastructure-as-codeAnsibleTerraformPrometheusGrafanaAWSGCPAzureContainerOrchestrationKubernetesDockerCI/CDJenkinsGitLab

Company Brief

Arista Networks
Designs and sells high-performance cloud networking hardware and software for large-scale data centers and enterprise environments, including switches, routers, and network operating systems focused on programmability, telemetry, and low-latency networking.
Industry: Networking Equipment
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 2004
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn