Senior Software Engineer, DGX Cloud Orchestration

NVIDIA
Santa Clara, United States
Workplace: HybridFull timeUSD 184,000 - 287,500 annuallyFunction: Software EngineeringSkills: ["Communication","Collaboration","Debugging","Problem-solving"]

Design and develop APIs and systems to orchestrate DGX Cloud operational workflows. Build state management and workflow automation to streamline infrastructure lifecycle processes, codify business processes into scalable systems, and ensure consistency. Integrate with Kubernetes and observability tools like Prometheus, OpenTelemetry, and Grafana, optimizing reliability and efficiency through telemetry-driven automation. Lead and ship technical projects with quality and scalability.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
2 days ago

Senior Software Engineer, DGX Cloud Orchestration

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 22 hours agoStatus: Live

Job Summary

Design and develop APIs and systems to orchestrate DGX Cloud operational workflows. Build state management and workflow automation to streamline infrastructure lifecycle processes, codify business processes into scalable systems, and ensure consistency. Integrate with Kubernetes and observability tools like Prometheus, OpenTelemetry, and Grafana, optimizing reliability and efficiency through telemetry-driven automation. Lead and ship technical projects with quality and scalability.
Location: Santa Clara, United States
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Design and develop APIs to orchestrate and integrate operational workflows.
  • •Build state management and workflow automation systems to streamline infrastructure lifecycle processes.
  • •Collaborate across teams to codify business processes into scalable, self-measuring systems.
  • •Develop extensible, schema-driven platforms to reduce manual toil and ensure operational consistency.
  • •Integrate with container orchestration and observability systems, optimizing reliability and efficiency using telemetry.

Pay and Benefits

Salary: USD 184,000 - 287,500 annually
Equity and Bonus:Equity

Key Requirements

  • •8+ years of industry experience with a Bachelors (or equivalent) degree; Master’s preferred.
  • •Experience designing, building, and operating services in a high-reliability environment.
  • •Proficiency in Go, Java, or Python.
  • •Strong understanding of cloud infrastructure (AWS, GCP, Azure) and container technologies like Docker and Kubernetes.
  • •Experience with high-scale distributed systems, including API and data pipeline architectural patterns.
Skills:CommunicationCollaborationDebuggingProblem-solving
Tech Stack:APIsGoJavaPythonAWSGCPAzureDockerKubernetesPrometheusOpenTelemetryGrafanaCUDACuDNN

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor