Sr. Site Reliability Engineer, AI Infrastructure (Starshield)
California, Redmond, Washington
Workplace: OnsiteFull timeUSD 165,000 - 265,000 annuallyFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Communications","Mentorship"]Design, operate, and scale on-prem infrastructure that powers AI clusters for Starshield’s national security missions. You’ll manage GPU/CPU deployments and provide GPU-as-a-service for customers, build automation for Kubernetes and operating systems, and own core platform components like databases, monitoring, and distributed storage. Partner closely with AI engineers to deliver highly scalable, operable services with strong availability, and mentor others while maintaining Top Secret clearance requirements.
Loading
Loading job details...
Preparing the role view and application actions.

