Forward Deployed Engineer: AI + HPC

Cedana
United States
Workplace: RemoteFull timeUSD 140,000 - 180,000 annuallyFunction: Data Science & Machine LearningExperience: 3-10 yearsSkills: ["Communication","Problem-solving","Teamwork"]

Forward Deployed Engineer at Cedana will lead end-to-end customer engagements deploying our AI+HPC platform across SLURM, Kubernetes, and NVIDIA Dynamo. You will own OS-level integrations, gather field feedback to drive product enhancements, and optimize reliability, throughput, and performance for enterprise and research customers, while ensuring scalable, secure deployments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cedana
Cedana
3 months ago

Forward Deployed Engineer: AI + HPC

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Forward Deployed Engineer at Cedana will lead end-to-end customer engagements deploying our AI+HPC platform across SLURM, Kubernetes, and NVIDIA Dynamo. You will own OS-level integrations, gather field feedback to drive product enhancements, and optimize reliability, throughput, and performance for enterprise and research customers, while ensuring scalable, secure deployments.
Location: United States
Workplace: Remote
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Entry level

Key Responsibilities

  • •Engineer solutions at client sites: Lead customer integrations and deploy Cedana into SLURM, Kubernetes, and Dynamo environments.
  • •Drive product innovation from the field: Identify technical gaps with clients and provide feedback that becomes core product features.
  • •Measure and optimize platform performance: Assess reliability and throughput and design policy-based migration automations to improve throughput and reliability.
  • •Own critical deployments: Ensure platform reliability for clients’ operations, debugging across the full stack and escalate when needed.
  • •Improve scalability: Create internal install playbooks to accelerate onboarding of subsequent customers.
Travel: Low travel

Pay and Benefits

Salary: USD 140,000 - 180,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVisionPaid Leave401k

Key Requirements

  • •3-10 years of software engineering experience with a track record of configuring and managing SLURM deployments.
  • •A multi-month enterprise or research deployment you led end-to-end, from scoping through signoff.
  • •Production experience standing up SLURM in a customer or research environment; configured slurmctld, slurmdbd, accounting, cgroup integration, and GPU resource selection.
  • •Strong Linux fundamentals of systemd, cgroups v2, namespaces, networking, filesystems, kernel module loading, PAM session modules.
  • •Working Kubernetes operations including operators, CRDs, device plugins, node-level debugging.
Experience:3-10 yearsAIHPCInfrastructure
Skills:CommunicationProblem-solvingTeamwork
Languages:English
Tech Stack:SLURMKubernetesNVIDIA DynamoLinuxSystemdCgroupsDmesgStrace

Eligibility

Visa:US citizen/visa only
Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Cedana
Develops AI-powered analytics and decision-support tools for enterprises, focusing on automating data ingestion, model-driven insights, and workflow optimization to help organizations extract actionable intelligence from complex datasets.
Industry: Data Infrastructure
Website