Principal Systems Software Engineer

Crusoe
San Francisco
Workplace: OnsiteFull timeUSD 260,000 - 340,000 annuallyFunction: Software EngineeringExperience: 12+ yearsEducation: mastersSkills: ["Communication","Leadership","Problem-solving","Opportunity finding","Technical strategy"]

Lead next-generation AI infrastructure by unifying Bare-Metal-as-a-Service, intelligent IaaS, and Elastic CaaS into a single high-performance compute fabric. Drive the technical roadmap for SR-IOV, RDMA, and virtualized GPU scheduling, and prototype/productionize advanced R&D in memory, networking, and compute. Serve as a hands-on, hyperscale-proven expert bridging silicon and software, authoring RFCs/white papers and resolving kernel-level race conditions.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Crusoe
Crusoe
1 day ago

Principal Systems Software Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Lead next-generation AI infrastructure by unifying Bare-Metal-as-a-Service, intelligent IaaS, and Elastic CaaS into a single high-performance compute fabric. Drive the technical roadmap for SR-IOV, RDMA, and virtualized GPU scheduling, and prototype/productionize advanced R&D in memory, networking, and compute. Serve as a hands-on, hyperscale-proven expert bridging silicon and software, authoring RFCs/white papers and resolving kernel-level race conditions.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Unify Bare-Metal-as-a-Service, Intelligent IaaS, and Elastic CaaS into a single high-performance pool of intelligence for AI workloads.
  • •Architect the internal cloud fabric, leveraging SR-IOV, RDMA, and virtualized GPU scheduling to drive the technical roadmap.
  • •Lead R&D workstreams to prototype and productionize novel approaches for managing memory, networking, and compute beyond standard cloud distributions.
  • •Provide hands-on kernel-level and orchestration leadership, including advanced debugging of race conditions and optimization of memory pinning for GPU clusters.
  • •Author white papers and RFCs and represent Crusoe in open-source communities and industry forums.

Pay and Benefits

Salary: USD 260,000 - 340,000 annually
Equity and Bonus:Equity
Perks:RsusHealth InsuranceDentalVision401kPaid Parental

Key Requirements

  • •12+ years of experience designing and shipping core infrastructure at a major hyperscaler (OCI, AWS, Azure, or GCP) or a specialized HPC cloud.
  • •Deep expertise with the Linux kernel, virtualization internals (KVM, QEMU, Firecracker), and high-performance networking (RoCE v2, InfiniBand).
  • •Ability to design software that maximizes performance of NVIDIA/AMD GPUs and high-speed NICs.
  • •Experience leading cross-functional teams through high-ambiguity projects and delivering production-ready, mission-critical systems.
  • •A Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related analytical field (or equivalent professional experience).
Experience:12+ yearsHyperscale infrastructureHPC cloud
Education:Master's in Computer Science, Computer Engineering, or a related analytical field
Skills:CommunicationLeadershipProblem-solvingOpportunity findingTechnical strategy
Tech Stack:Linux kernelKVMQEMUFirecrackerKubernetesSlurmSR-IOVRDMAInfiniBandRoCE v2NVIDIA GPUsAMD GPUsVirtualized GPU schedulingMemory pinning

Company Brief

Crusoe
Builds vertically integrated, energy-first AI infrastructure and purpose-built AI data centers (Crusoe Cloud), leveraging clean/stranded energy to power large-scale GPU compute for AI training and inference.
Industry: Data Centers
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Decacorn (USD 10B+)
Funding: Series E+
Headquarters: Denver, United States
Founded: 2018
Glassdoor
Glassdoor: 3.7
WebsiteLinkedInGlassdoor