Solutions Architect - DevOps

NVIDIA
Australia
Workplace: RemoteFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 5+ yearsEducation: mastersSkills: ["Communication","Collaboration","Leadership","Problem-solving","Customer engagement"]

Senior Cloud Infrastructure and DevOps Solutions Architect to join the ANZ team, focusing on stand-up and operational excellence for Neo Clouds. Works with customers, partners and internal teams to design, implement and operate large-scale AI/HPC infrastructure, Kubernetes-based platforms, automation, and hardware networking. Ideal location may be Sydney or Melbourne; involves leadership on DevOps and platform architecture for enterprise, research and academic environments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
4 months ago

Solutions Architect - DevOps

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Senior Cloud Infrastructure and DevOps Solutions Architect to join the ANZ team, focusing on stand-up and operational excellence for Neo Clouds. Works with customers, partners and internal teams to design, implement and operate large-scale AI/HPC infrastructure, Kubernetes-based platforms, automation, and hardware networking. Ideal location may be Sydney or Melbourne; involves leadership on DevOps and platform architecture for enterprise, research and academic environments.
Location: Australia
Workplace: Remote
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Maintain large-scale computational and AI infrastructure, focusing on monitoring, logging, workload orchestration (Kubernetes and Linux job schedulers).
  • •Optimize scalable, production-ready Kubernetes-based container platforms coordinated with enterprise-grade networking and storage.
  • •Serve as a key technical resource, develop, refine, and document standard methodologies and operational guidelines to be shared with internal teams.
  • •Perform end-to-end resolution across the stack, from bare metal and OS through the software stack, container platform, networking, and storage.
  • •Develop tooling to automate deployment and management of large-scale infrastructure environments, to enable self-service consumption of resources.

Key Requirements

  • •BS/MS/PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields, with 5+ years of professional experience in managing scalable cloud environments and automation engineering roles.
  • •Extensive experience with Kubernetes for container orchestration, resource scheduling, scaling, and integration with HPC environments.
  • •Proven understanding of networking fundamentals (TCP/IP stack), data center architectures, and hands-on experience managing HPC/AI clusters, including deployment, optimization, and fixing issues.
  • •Familiarity with HPC and AI technologies (CPUs, GPUs, high-speed interconnects) and supporting software stacks.
  • •Deep knowledge of Linux (RedHat/CentOS, Ubuntu), OS-level security, and protocols (TCP, DHCP, DNS). Experience with storage solutions such as Lustre, GPFS, ZFS, XFS, and emerging Kubernetes storage technologies.
  • •Proficiency in Python and Bash scripting, configuration management, and Infrastructure-as-Code tools (e.g., Ansible, Terraform). Experience with observability stacks (Grafana, Loki, Prometheus).
  • •Strong background in crafting scalable solutions and providing consultative support to customers.
Experience:5+ yearsCloudAIHPCKubernetesData centerInfrastructure
Education:Master's
Skills:CommunicationCollaborationLeadershipProblem-solvingCustomer engagement
Tech Stack:KubernetesLinuxPythonBashAnsibleTerraformGrafanaPrometheusLokiLustreGPFSZFSXFSInfiniBandRoCECUDANVIDIA DGX

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor