Site Reliability Operations Engineer

Salesforce
Seattle
Workplace: OnsiteFull timeUSD 94,000 - 142,300 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 5-8 yearsSkills: ["Incident management","Problem solving","Communication","Troubleshooting","Stakeholder management"]

Own incident response and reliability operations for Salesforce’s internal DET environment. Serve as Incident Commander, troubleshoot enterprise infrastructure, applications, and networking across multiple platforms, and coordinate emergency changes to restore service quickly. Improve incident response through runbooks, SOPs, automation, and problem management, analyze incident data and KPIs, and participate in regional on-call coverage to reduce toil and improve performance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Salesforce
Salesforce
1 day ago

Site Reliability Operations Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Own incident response and reliability operations for Salesforce’s internal DET environment. Serve as Incident Commander, troubleshoot enterprise infrastructure, applications, and networking across multiple platforms, and coordinate emergency changes to restore service quickly. Improve incident response through runbooks, SOPs, automation, and problem management, analyze incident data and KPIs, and participate in regional on-call coverage to reduce toil and improve performance.
Location: Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Respond to and manage major incidents, serving as Incident Commander to coordinate teams and drive rapid service restoration.
  • •Monitor and troubleshoot enterprise systems across infrastructure, applications, and network components.
  • •Improve incident response by creating/improving runbooks and SOPs and driving automation.
  • •Coordinate emergency changes and infrastructure updates to resolve incidents and maintain business continuity.
  • •Lead problem management, investigate recurring incidents, document root cause analyses, and track known errors.

Pay and Benefits

Salary: USD 94,000 - 142,300 annually
Perks:MedicalDentalVisionHealth InsurancePaid ParentalLife InsuranceDisability Insurance401kEquityTime Off

Key Requirements

  • •5-8 years in IT operations, incident management, or site reliability work.
  • •Experience supporting 24x7 high-availability environments with enterprise systems preferred.
  • •Demonstrated ability to manage high-severity incidents under pressure and make decisions balancing technical and business needs.
  • •Strong verbal and written communication skills for technical and executive audiences, including incident updates and status reports.
  • •A related technical degree is required.
Experience:5-8 yearsEnterprise IT24x7 high availability
Education:
Skills:Incident managementProblem solvingCommunicationTroubleshootingStakeholder management
Certifications:ITILAWSCCNAMCSARHCE
Tech Stack:AWSWindowsLinuxPythonBashPowerShellSplunkGrafanaTableauPuppetChefITILNetworkingCloud platformsVirtualization technologies

Company Brief

Salesforce
Provides a leading cloud-based customer relationship management (CRM) platform with sales, service, marketing, analytics, and integration tools that empower businesses to manage customer relationships and digital transformation at scale.
Industry: SaaS
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Francisco, United States
Founded: 1999
Glassdoor
Glassdoor: 4.1
WebsiteLinkedInGlassdoor