Staff AI Systems Engineer

Archer
San Jose
Workplace: OnsiteFull timeUSD 172,800 - 216,000 annuallyFunction: IT Operations (Systems/Network Admin)Experience: 5+ yearsSkills: ["Cross-functional collaboration","Debugging","Performance optimization"]

Architect, deploy, and manage infrastructure services that enable large-scale AI training and low-latency inference. Build and maintain MLOps tooling and end-to-end workflows using MLflow, while optimizing compute across multi-cloud environments and GPU/AI clouds. Collaborate with AI researchers and software engineers to productionize models, implement monitoring, and troubleshoot performance bottlenecks at the hardware-software boundary.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Archer
Archer
2 months ago

Staff AI Systems Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 22 hours agoStatus: Live

Job Summary

Architect, deploy, and manage infrastructure services that enable large-scale AI training and low-latency inference. Build and maintain MLOps tooling and end-to-end workflows using MLflow, while optimizing compute across multi-cloud environments and GPU/AI clouds. Collaborate with AI researchers and software engineers to productionize models, implement monitoring, and troubleshoot performance bottlenecks at the hardware-software boundary.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: IT Operations (Systems/Network Admin)
Seniority: Mid level

Key Responsibilities

  • •Architect, deploy, and manage infrastructure services for large-scale AI training and low-latency inference.
  • •Deploy and scale resilient infrastructure for distributed AI model training and implement necessary orchestration.
  • •Utilize and maintain end-to-end MLOps tooling, including MLflow for experiment tracking and model registry.
  • •Optimize compute and inference by maximizing hardware utilization and managing multi-cloud scheduling and LLM serving engines.
  • •Partner with AI researchers and software engineers to productionize models, establish monitoring, and debug performance bottlenecks.

Pay and Benefits

Salary: USD 172,800 - 216,000 annually

Key Requirements

  • •BS/MS/PhD degree in Computer Science, Software Engineering, or a related field.
  • •5+ years of professional software engineering experience focused on AI/ML systems, high-performance computing (HPC), or ML infrastructure.
  • •Familiarity with AWS hyper-scaler infrastructure and specialized AI-centric bare-metal/GPU clouds such as Nebius AI Cloud.
  • •Hands-on experience with Docker and Kubernetes, plus cloud-agnostic cluster abstraction (e.g., SkyPilot) for multi-region GPU availability.
  • •Deep understanding of LLM serving at scale using frameworks like vLLM and SGLang, and experience building high-throughput data pipelines (SQL, NoSQL, Parquet).
Experience:5+ yearsAI/MLHigh-performance computing (HPC)
Skills:Cross-functional collaborationDebuggingPerformance optimization
Languages:English
Tech Stack:AWSNebius AI CloudMLflowDockerKubernetesSkyPilotVLLMSGLangSQLNoSQLParquetHPCLLM serving engines

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Archer
Develops electric vertical takeoff and landing (eVTOL) aircraft and urban air mobility systems to enable sustainable, intra-city air transportation for passengers and cargo.
Industry: Aerospace Manufacturing
Company Size: Large (251 to 1,000 employees)
Revenue: Pre-Revenue (USD 0)
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Palo Alto, United States
Founded: 2018
WebsiteLinkedIn