Site Reliability Engineering Technical Leader

Cisco
Irvine, California, Colorado, Arizona, Oregon, Texas
Workplace: RemoteFull timeUSD 192,400 - 275,800 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 12+ yearsSkills: ["Technical leadership","Communication","Mentoring","Technical judgment","Root cause analysis"]

Serve as the most senior technical individual contributor for the CloudOps team, owning incident response strategy and high-stakes P1/P2 escalations to protect Splunk Cloud uptime. Lead technical ownership of complex customer stacks, drive automation and infrastructure architecture decisions, and shape operational processes through post-mortems and systemic improvements. Calibrate and raise engineering judgment by building enduring frameworks used across the organization.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cisco
Cisco
14 hours ago

Site Reliability Engineering Technical Leader

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Serve as the most senior technical individual contributor for the CloudOps team, owning incident response strategy and high-stakes P1/P2 escalations to protect Splunk Cloud uptime. Lead technical ownership of complex customer stacks, drive automation and infrastructure architecture decisions, and shape operational processes through post-mortems and systemic improvements. Calibrate and raise engineering judgment by building enduring frameworks used across the organization.
Location: Irvine, California, Colorado, Arizona, Oregon, Texas
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Set technical judgment, architectural thinking, and operational excellence across the organization as the team’s most senior IC.
  • •Own the technical strategy for how the team responds to, learns from, and prevents critical incidents, serving as the escalation point during P1/P2 events.
  • •Hold deep ownership of the most complex customer stacks and shape automation direction and infrastructure architecture decisions.
  • •Influence how Splunk Cloud operational processes evolve and advise Engineering, Customer Success, and Release Management during demanding customer situations.
  • •Raise the technical floor by calibrating senior engineers and building frameworks that persist beyond individual incidents.

Pay and Benefits

Salary: USD 192,400 - 275,800 annually
Perks:Health InsuranceDentalVision401kPaid ParentalLong-term DisabilityLife InsurancePaid HolidaysPaid LeaveSick TimeRsus

Key Requirements

  • •Bachelors + 12 years of related experience, or Masters + 8 years, or PhD + 5 years.
  • •7+ years in SRE, cloud operations, and systems engineering with Linux administration.
  • •6+ years hands-on experience across AWS, GCP, or Azure.
  • •5+ years experience in scripting/automation such as Python or Go/golang.
  • •4+ years leading post-mortems and root cause analysis for high-severity events to drive systemic improvements.
Experience:12+ yearsSRECloud operationsSystems engineeringLinuxAWSGCPAzureEnterprise scale
Skills:Technical leadershipCommunicationMentoringTechnical judgmentRoot cause analysis
Tech Stack:LinuxAWSGCPAzurePythonGoGolangSplunkSPLIndexer clusteringSearch Head ClustersKVStoreSplunk Observability

Company Brief

Cisco
Global technology company that designs, manufactures, and sells networking hardware, telecommunications equipment, and high-technology services and products for enterprises, service providers, and governments worldwide.
Industry: Networking Equipment
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 1984
Glassdoor
Glassdoor: 4.0
WebsiteLinkedInGlassdoor