Software Engineer, Infrastructure

Exa
Singapore
Workplace: OnsiteFull timeSGD 90,000 - 300,000 annuallyFunction: Software EngineeringSkills: ["Kubernetes","Rust","GPU","Observability","Scheduling","Batch","Ray"]

Join Exa’s Infrastructure Team to design and operate large-scale AI infrastructure, including GPU clusters, Kubernetes, and cloud batch systems. You’ll help build the software that schedules, monitors, and optimizes a multi-thousand-node cluster powering state-of-the-art embedding models. Expect hands-on work across the stack, from orchestration to observability, in a fast-moving, in-person Singapore office.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Exa
Exa
8 months ago

Software Engineer, Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Join Exa’s Infrastructure Team to design and operate large-scale AI infrastructure, including GPU clusters, Kubernetes, and cloud batch systems. You’ll help build the software that schedules, monitors, and optimizes a multi-thousand-node cluster powering state-of-the-art embedding models. Expect hands-on work across the stack, from orchestration to observability, in a fast-moving, in-person Singapore office.
Location: Singapore
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Design and operate large-scale infrastructure (GPU clusters, Kubernetes clusters, or cloud batchjob systems).
  • •Contribute to building GPU cluster orchestration and instrumentation to maximize reliability and efficiency.
  • •Develop and improve tooling for deployment, monitoring, and observability across the stack.
  • •Collaborate with the infra and platform teams to scale systems as usage grows.
  • •Focus on optimization and performance across the entire infrastructure stack.

Pay and Benefits

Salary: SGD 90,000 - 300,000 annually
Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •Experience designing and operating large-scale infrastructure—GPU clusters or large Kubernetes clusters or cloud batchjob systems.
  • •Strong focus on reliability, observability, and optimization across the entire stack.
  • •Ability to design and implement tooling for deployment, monitoring, and performance tuning.
  • •Proficiency with container orchestration (Kubernetes) and distributed systems concepts.
  • •Experience with GPU scheduling or high-performance computing is a plus.
Experience:InfrastructureCloud
Skills:KubernetesRustGPUObservabilitySchedulingBatchRay
Tech Stack:KubernetesRayRustGPUAWSObservabilityBatch processing

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Exa
Builds an embeddings-based neural search engine and web search API for AI applications, offering crawling, retrieval, and deep research tools to serve developers and AI agents with up-to-date web knowledge.
Industry: API Platforms
Company Size: Small (11 to 50 employees)
Growth: Scaleup
Valuation: USD 500M to 1B
Funding: Series B
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn