Senior System Software Engineer - Scientific Computing PaaS

NVIDIA
Santa Clara, United States
Full timeUSD 184,000 - 356,500 annuallyFunction: Software EngineeringEducation: bachelorsSkills: ["Interpersonal skills","Independent work"]

Build and operate cloud-based scientific computing platform workflows, powering physics-informed and data-driven applications across simulation and AI training/inference. Own core cloud services and infrastructure from system design through deployment and operations, including scalable I/O, checkpointing, and data pipelines. Optimize compute, storage, and network architecture for high-performance workloads using CPUs/GPUs, distributed systems, parallel processing, and large-scale debugging to improve reliability and performance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
3 days ago

Senior System Software Engineer - Scientific Computing PaaS

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Build and operate cloud-based scientific computing platform workflows, powering physics-informed and data-driven applications across simulation and AI training/inference. Own core cloud services and infrastructure from system design through deployment and operations, including scalable I/O, checkpointing, and data pipelines. Optimize compute, storage, and network architecture for high-performance workloads using CPUs/GPUs, distributed systems, parallel processing, and large-scale debugging to improve reliability and performance.
Location: Santa Clara, United States
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Design services and take ownership of the underlying cloud infrastructure for physics-informed and data-driven scientific workflows.
  • •Design novel algorithms and partner with operations to improve end-to-end system performance across the stack.
  • •Design, build, deploy, and operate scalable I/O infrastructure for checkpointing and data loading plus pre-/post-processing.
  • •Optimize compute, storage, and network architecture for physics and simulation workloads.
  • •Debug and improve reliability/performance across processes, threads, synchronization, scheduling, IPC, memory management, and filesystem/I-O structure.

Pay and Benefits

Salary: USD 184,000 - 356,500 annually
Equity and Bonus:Equity

Key Requirements

  • •BS/MS in Computer Science (or related) or equivalent experience.
  • •10+ years building and operating distributed compute and data-intensive platforms as a service on cloud.
  • •Proven skill in a compiled language such as Go, Rust, or C++ (or equivalent).
  • •Strong foundational knowledge of cloud computing concepts including datacenter architecture, cloud security architecture, virtualization, and resource pooling/elasticity.
  • •Strong distributed systems and parallel processing skills, including synchronization, fault tolerance/failure detection, consensus protocols, and cluster scalability/performance.
Education:Bachelor's
Skills:Interpersonal skillsIndependent work
Tech Stack:GoRustC++MicroservicesAPIsGPUsCPUsMPINCCLDistributed systemsParallel processingCheckpointingOrchestration

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor