Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco, Seattle, New York
Workplace: HybridFull timeUSD 250,000 - 485,000 annuallyFunction: Software EngineeringSkills: ["Ownership","End-to-end problem-solving","Technical partnership","Reliability focus","Distributed systems thinking"]Own the platform that powers Perplexity’s real-time training and inference workloads on a multi-cloud GPU fleet. Build a self-serve compute platform, operate GPU provisioning and lifecycle management, and design scheduling/placement to address GPU scarcity. Develop Kubernetes-based GPU orchestration (operators, CRDs, multi-cluster federation) with fault tolerance, autoscaling, and observability so both long-running training and low-latency inference run reliably. Partner with inference and cloud engineers to set the technical roadmap.
Loading
Loading job details...
Preparing the role view and application actions.

