Staff Site Reliability Engineer

Attentive
United States
Workplace: OnsiteFull timeUSD 180,000 - 240,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 7+ yearsSkills: ["Problem-solving","Strategic thinking","Mentorship","Communication","Technical leadership","Cross-team collaboration"]

Design and implement production systems that improve reliability, observability, traceability, and incident management for a platform processing billions of events daily. Lead cross-team initiatives and technical collaborations across AI/ML, Data, Platform, and Product. Define production standards and reliability metrics (SLIs/SLOs), drive cost optimization, and mentor others while influencing technical roadmaps to scale safely and securely.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Attentive
Attentive
1 day ago

Staff Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Design and implement production systems that improve reliability, observability, traceability, and incident management for a platform processing billions of events daily. Lead cross-team initiatives and technical collaborations across AI/ML, Data, Platform, and Product. Define production standards and reliability metrics (SLIs/SLOs), drive cost optimization, and mentor others while influencing technical roadmaps to scale safely and securely.
Location: United States
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Sr. Manager level

Key Responsibilities

  • •Design and implement systems that enhance reliability, observability, traceability, and incident management so the platform scales effectively
  • •Lead cross-team strategic initiatives by providing technical leadership and guidance
  • •Collaborate with engineers across AI/ML, Data, Platform, and Product to develop best-in-class services
  • •Define and enforce production standards, processes, and tools to ensure operational excellence
  • •Advocate for and implement SLIs, SLOs, and other reliability-focused metrics while mentoring team members

Pay and Benefits

Salary: USD 180,000 - 240,000 annually
Perks:Health InsuranceEquityBenefits

Key Requirements

  • •7+ years of experience in Production Engineering, Backend Engineering, SRE, DevOps, or similar roles
  • •Strong coding ability in at least one language (e.g., Golang, Python, Java, TypeScript) to solve complex issues
  • •Demonstrated success delivering medium to large-scale projects improving platform reliability and scalability
  • •Deep understanding of production reliability concepts, including SLIs, SLOs, and incident management
  • •Excellent verbal and written communication skills to influence and collaborate across technical and non-technical teams
Experience:7+ yearsProduction engineeringBackend engineeringSREDevOpsPlatform reliabilityObservability
Skills:Problem-solvingStrategic thinkingMentorshipCommunicationTechnical leadershipCross-team collaboration
Languages:English
Tech Stack:GolangPythonJavaTypeScriptSLIsSLOs

Company Brief

Attentive
Provides a mobile messaging and personalization platform that helps brands engage customers via SMS, email, and onsite messaging to drive sales, retention, and personalized marketing at scale.
Industry: Enterprise Software
Company Size: Enterprise (1,001+ employees)
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2016
WebsiteLinkedIn