Technical Marketing Engineer – Datacenter GPU Clusters & Networking

AMD
Santa Clara
Workplace: OnsiteFull timeFunction: Marketing & GrowthEducation: bachelorsSkills: ["Communication","Presentation","Technical writing","Collaboration","Problem-solving"]

Technical Marketing Engineer bridging AI training performance with clear technical storytelling. You will develop performance-focused content, act as a subject matter expert on AI training across the full model lifecycle, and enable customers and internal teams to optimize AMD Instinct GPUs for large-scale AI workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
4 months ago

Technical Marketing Engineer – Datacenter GPU Clusters & Networking

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 hours agoStatus: Live

Job Summary

Technical Marketing Engineer bridging AI training performance with clear technical storytelling. You will develop performance-focused content, act as a subject matter expert on AI training across the full model lifecycle, and enable customers and internal teams to optimize AMD Instinct GPUs for large-scale AI workloads.
Location: Santa Clara
Workplace: Onsite
Employment Type: Full time
Job Function: Marketing & Growth

Key Responsibilities

  • •Partner with AMD’s AI software engineering team to develop performance-focused technical content for AI training workloads, including optimization guides, benchmarking results, scaling studies, and tuning methodologies.
  • •Serve as a subject matter expert on AI training performance across the full model lifecycle, including large-scale pre-training, fine-tuning, RL workflows, distillation, and QAT.
  • •Develop and publish deep technical content for training workloads, including performance analyses, scaling studies, optimization guides, and distributed training best practices.
  • •Analyze and optimize training performance across compute utilization, memory efficiency, communication overhead, and scaling behavior in distributed environments.
  • •Engage with internal and external experts to validate performance claims, and translate low-level optimizations into actionable guidance for product improvements.

Key Requirements

  • •5+ years optimizing AI training workloads on GPUs or accelerators at scale.
  • •Hands-on experience with distributed training frameworks (e.g., PyTorch, DeepSpeed, Megatron-LM).
  • •Strong programming skills in Python and/or C/C++.
  • •Experience with ROCm, CUDA, or similar GPU compute stacks.
  • •Experience working with ISVs, hyperscalers, or large-scale AI deployments.
Experience:AIGPUData centersMachine learningDeep learning
Education:Bachelor's in Computer Science
Skills:CommunicationPresentationTechnical writingCollaborationProblem-solving
Languages:English
Tech Stack:PythonC++CUDAROCmPyTorchDeepSpeedMegatron-LMMarkdownRead the DocsJupyter Notebooks

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn