Senior Deep Learning Software Engineer, Inference

NVIDIA
Netherlands, Poland, United Kingdom
Workplace: RemoteFull timePLN 221,250 - 383,500 annuallyFunction: Software EngineeringExperience: 5+ yearsEducation: phdSkills: ["Performance optimization","Performance analysis","Tuning","Software design","Cross-collaboration","Debugging","Profiling","Code optimization"]

Design, build, and optimize GPU-accelerated deep learning inference software for large-scale model serving. Help improve high-performance frameworks such as SGLang and vLLM, contributing performance analysis, tuning, and new features for NVIDIA inference libraries and related solutions. Work across CPU/GPU architectures to scale LLM and generative AI workloads across NVIDIA accelerators, using tools like CUDA, Triton, CUTLASS, and NCCL to deliver efficient deployments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
22 hours ago

Senior Deep Learning Software Engineer, Inference

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 22 hours agoStatus: Live

Job Summary

Design, build, and optimize GPU-accelerated deep learning inference software for large-scale model serving. Help improve high-performance frameworks such as SGLang and vLLM, contributing performance analysis, tuning, and new features for NVIDIA inference libraries and related solutions. Work across CPU/GPU architectures to scale LLM and generative AI workloads across NVIDIA accelerators, using tools like CUDA, Triton, CUTLASS, and NCCL to deliver efficient deployments.
Location: Netherlands, Poland, United Kingdom
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Design, build, and optimize GPU-accelerated software powering efficient deep learning inference and model serving.
  • •Perform performance optimization, analysis, and tuning of deep learning models for LLM, multimodal, and generative AI domains.
  • •Scale deep learning model performance across NVIDIA accelerator architectures and systems.
  • •Contribute features and code to NVIDIA inference libraries and solutions including vLLM and SGLang and related software offerings.
  • •Collaborate with cross-functional teams across frameworks and NVIDIA libraries to deliver inference optimization solutions.

Pay and Benefits

Salary: PLN 221,250 - 383,500 annually

Key Requirements

  • •Masters or PhD or equivalent experience in Computer Engineering, Computer Science, EECS, or AI.
  • •5+ years of relevant software development experience.
  • •Excellent C/C++ programming and software design skills; SW Agile skills helpful and Python experience a plus.
  • •Prior experience training, deploying, or optimizing inference of deep learning models in production is a plus.
  • •Prior experience with performance modeling, profiling, debugging, and code optimization or architectural knowledge of CPU/GPU is a plus.
Experience:5+ yearsDeep learningLLMGenerative AIProduction deploymentGPU accelerationHigh-performance computingInference optimization
Education:PhD / Doctorate
Skills:Performance optimizationPerformance analysisTuningSoftware designCross-collaborationDebuggingProfilingCode optimization
Tech Stack:CC++PythonSGLangVLLMFlashInferCUDACUDA kernelsOAI TritonCUTLASSNCCLNVSHMEMPyTorchNVIDIA acceleratorsGPU programmingCPU and GPU architectureMulti-GPU communicationsLLMMultimodalGenerative AI

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor