Site Reliability Engineer, AI Infrastructure (Starshield)
Washington, California, Redmond
Workplace: OnsiteFull timeUSD 125,000 - 195,000Function: DevOps, Cloud & InfrastructureExperience: 1+ yearsEducation: bachelorsSkills: ["Communication"]Design, operate, and scale on-premise infrastructure that powers Starshield’s AI and GPU workloads for critical national security missions. Build automation for deployments and manage Kubernetes/AI clusters, GPU-as-a-service on bare metal and virtualized platforms, and core services like databases, monitoring, and distributed storage. Partner with AI engineers to productize highly scalable, maintainable systems with strong availability, lifecycle ownership, and continuous monitoring.
Loading
Loading job details...
Preparing the role view and application actions.

