Senior Software Engineer, Core Infrastructure Services - DGX Cloud

NVIDIA
Santa Clara, Texas, Colorado, California, Massachusetts
Workplace: RemoteFull timeUSD 168,000 - 322,000 annuallyFunction: Software EngineeringExperience: 8+ yearsEducation: bachelorsSkills: ["Problem-solving","Communication","Collaboration"]

Build and operate core infrastructure services that power NVIDIA’s DGX Cloud and SuperPod deployments. Design and develop secure, scalable cloud-native platform services, including orchestration, self-service workflows, and automation. Own integrations for infrastructure provisioning and lifecycle management, while improving reliability through observability and security. Partner with infrastructure and networking teams to deliver production-scale services at global scale.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Software Engineer, Core Infrastructure Services - DGX Cloud

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Build and operate core infrastructure services that power NVIDIA’s DGX Cloud and SuperPod deployments. Design and develop secure, scalable cloud-native platform services, including orchestration, self-service workflows, and automation. Own integrations for infrastructure provisioning and lifecycle management, while improving reliability through observability and security. Partner with infrastructure and networking teams to deliver production-scale services at global scale.
Location: Santa Clara, Texas, Colorado, California, Massachusetts
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Build and operate core infrastructure services that power NVIDIA’s global AI infrastructure.
  • •Architect and develop secure, scalable, highly available cloud-native platform services.
  • •Develop software for infrastructure orchestration, self-service workflows, and platform automation.
  • •Own integrations that automate infrastructure provisioning and lifecycle management.
  • •Improve reliability and resilience through observability and security, and drive operational excellence via automation, monitoring, incident response, and continuous improvement.

Pay and Benefits

Salary: USD 168,000 - 322,000 annually
Equity and Bonus:Equity

Key Requirements

  • •BS or equivalent experience with 8+ years of relevant industry experience.
  • •Strong proficiency in Python and Go, building production-quality software.
  • •Experience with cloud-native microservices and APIs on Kubernetes using FastAPI, gRPC, or REST.
  • •Experience with infrastructure automation (Terraform, Ansible), workflow orchestration (Temporal), and distributed systems (Redis, Kafka, NATS, SQS).
  • •Experience designing, building, and operating production infrastructure services and observability/security capabilities (DNS, NTP, Prometheus, Grafana, OpenTelemetry, VPNs, firewalls).
Experience:8+ yearsCloud-nativeMicroservicesDistributed systemsInfrastructure automationPublic cloud
Education:Bachelor's
Skills:Problem-solvingCommunicationCollaboration
Tech Stack:PythonGoKubernetesFastAPIGRPCRESTTerraformAnsibleTemporalRedisKafkaNATSSQSDNSNTPRADIUSOAuthPrometheusGrafanaOpenTelemetry

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor