Senior System Software Engineer - AI Data Platform - Inference Factory Optimization

NVIDIA
Hanoi, Vietnam
Workplace: HybridFull timeFunction: Software EngineeringExperience: 5+ yearsEducation: bachelorsSkills: ["Problem-solving","Analytical thinking","Debugging","Collaboration","Communication"]

Build and optimize scalable automation systems for validating, tuning, and deploying AI models across cloud and on-prem data centers. Develop infrastructure and tools for performance optimization using test harnesses, benchmarking, and analytical frameworks. Apply deep system software expertise—OS/kernel internals, device drivers, memory management, storage, networking, and high-speed interconnects—while partnering with engineering teams to define requirements, deliver solutions, and improve system reliability.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 month ago

Senior System Software Engineer - AI Data Platform - Inference Factory Optimization

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 21 hours agoStatus: Live

Job Summary

Build and optimize scalable automation systems for validating, tuning, and deploying AI models across cloud and on-prem data centers. Develop infrastructure and tools for performance optimization using test harnesses, benchmarking, and analytical frameworks. Apply deep system software expertise—OS/kernel internals, device drivers, memory management, storage, networking, and high-speed interconnects—while partnering with engineering teams to define requirements, deliver solutions, and improve system reliability.
Location: Hanoi, Vietnam
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Develop infrastructure and tools to automate complex software processes.
  • •Implement test harnesses, benchmarking frameworks, and analytical tools to optimize performance and efficiency of software and hardware platforms.
  • •Use expertise in OS/kernel internals, device drivers, memory management, storage, networking, and high-speed interconnects to build and troubleshoot performant systems.
  • •Partner with engineering teams to understand needs, define requirements, and deliver efficient solutions.
  • •Set performance goals, monitor feedback, analyze data, and continuously improve system reliability and technical roadmaps for platform automation.

Key Requirements

  • •Bachelor's or equivalent experience in Computer Science/Computer Engineering (or a Master’s degree or equivalent in a similar field).
  • •5+ years of software development experience focused on infrastructure, distributed systems, automation, and/or performance engineering.
  • •System-level programming expertise using C++, Python, or Go for building tools and automation.
  • •Deep understanding of system software, including operating system internals, device drivers, memory management, and performance debugging.
  • •Experience designing and operating large-scale distributed systems, including networking protocols, cluster management, and high-performance interconnects.
Experience:5+ years
Education:Bachelor's
Skills:Problem-solvingAnalytical thinkingDebuggingCollaborationCommunication
Tech Stack:C++PythonGoDockerKubernetes

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor