Sr. Staff Software Development Engineer - Collectives and Network optimization

AMD
San Jose
Workplace: HybridFull timeFunction: Product ManagementEducation: mastersSkills: ["Leadership","Collaboration","Communication","Mentoring","Presentation"]

Senior-level engineer responsible for driving AMD’s strategy, architecture, optimization and tooling to achieve industry-leading AI pre-training and distributed inference performance on AMD GPUs. Partner across hardware architecture, AI frameworks, compilers, runtime, ROCm and developer tools to scale performance analysis and optimization for Collectives and Network, optimizing across multiple generations of AMD GPUs and state-of-the-art AI models.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
2 months ago

Sr. Staff Software Development Engineer - Collectives and Network optimization

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Senior-level engineer responsible for driving AMD’s strategy, architecture, optimization and tooling to achieve industry-leading AI pre-training and distributed inference performance on AMD GPUs. Partner across hardware architecture, AI frameworks, compilers, runtime, ROCm and developer tools to scale performance analysis and optimization for Collectives and Network, optimizing across multiple generations of AMD GPUs and state-of-the-art AI models.
Location: San Jose
Workplace: Hybrid
Employment Type: Full time
Job Function: Product Management

Key Responsibilities

  • •Help set strategy and roadmap for AMD Collectives and Network optimizations.
  • •Provide guidelines to customers on efficient network load-balancing, workload scheduling and model sharding strategies.
  • •Performance tuning, profiling and analysis of large-scale models for LLM, diffusion, multimodal, RecSys and generative AI, single node and distributed.
  • •Participate in hardware-software co-design for future hardware optimizations – especially on scale-up networks, NIC and scale-out networks.
  • •Develop and improve framework, tools and infrastructure for performance estimation, modeling and reporting.

Key Requirements

  • •Strong background in network, NIC and GPU hardware architecture with hands-on performance optimization experience.
  • •Experience with AI frameworks such as PyTorch, JAX, vLLM, and SGLang.
  • •Proven ability to map model architectures to low-level software and hardware, and optimize distributed inference.
  • •Experience with performance modeling, profiling large-scale models (LLMs, diffusion, multimodal).
  • •Excellent communication and cross-functional collaboration skills; leadership/mentoring in a research/engineering team.
Experience:AIGPUDistributed systemsNetwork optimization
Education:Master's in Computer Science
Skills:LeadershipCollaborationCommunicationMentoringPresentation
Languages:English
Tech Stack:PythonPyTorchJAXVLLMSGLangROCmNICGPU

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn