Staff Site Reliability Engineer - AI Platform Runtime
Santa Clara
Workplace: HybridFull timeUSD 168,000 - 333,500 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsEducation: bachelorsSkills: ["Problem-solving","Communication","Teamwork","Technical leadership","Ownership","Mentorship"]Lead SRE technical strategy for large-scale, cross-functional initiatives that improve reliability, scalability, and developer productivity. Design and build resilient distributed systems powering next-generation AI-driven enterprise products. Drive automation and observability improvements using metrics and analytics, and collaborate across Cloud, Platform, Security, and AI/ML teams to ensure high availability and secure operations. Mentor engineers and champion best practices in incident management and postmortem analysis.
Loading
Loading job details...
Preparing the role view and application actions.

