Lead Performance and Optimization Engineer

AMD
Bengaluru
Workplace: OnsiteFull timeINR 3,245,970 - 4,637,100 annuallyFunction: Administration & Executive AssistanceExperience: 8-12 yearsEducation: bachelorsSkills: ["Python","Bash","VTune","Ftrace","ZenDNN","PyTorch","TensorFlow","ONNX Runtime","EPYC","NUMA","HugePages","Perf"]

Lead performance engineering for CPU inference workloads on AMD EPYC platforms, benchmarking and optimizing ML inference using ZenDNN with PyTorch/TensorFlow/ONNX Runtime. Hands-on OS/BIOS tuning, thread/core affinity, NUMA/memory optimization, and automation to improve throughput/latency across single and multi-instance scenarios.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
3 months ago

Lead Performance and Optimization Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 18 hours agoStatus: Live

Job Summary

Lead performance engineering for CPU inference workloads on AMD EPYC platforms, benchmarking and optimizing ML inference using ZenDNN with PyTorch/TensorFlow/ONNX Runtime. Hands-on OS/BIOS tuning, thread/core affinity, NUMA/memory optimization, and automation to improve throughput/latency across single and multi-instance scenarios.
Location: Bengaluru
Workplace: Onsite
Employment Type: Full time
Job Function: Administration & Executive Assistance
Seniority: Manager level

Key Responsibilities

  • •Lead performance engineering and optimization for CPU inference workloads, focusing on throughput and latency improvements across single and multi-instance executions.
  • •Configure and tune OS/BIOS parameters (NUMA, HugePages, SMT, power/performance modes) to optimize CPU performance.
  • •Design and validate thread/core affinity strategies, CPU isolation, scheduling, and multi-instance execution models.
  • •Collaborate with ZenDNN and kernel/ops teams to drive performance improvements and guide kernel-level optimizations.
  • •Develop automation pipelines and Python/Bash tooling to manage experiment workflows, data capture, and performance reporting.

Pay and Benefits

Salary: INR 3,245,970 - 4,637,100 annually

Key Requirements

  • •8-12 years in performance engineering, systems optimization, or low-level Linux execution
  • •Strong expertise in server-class CPUs, CPU microarchitecture, and ML inference
  • •Proficiency with EPYC-optimized ML libraries (ZenDNN) and frameworks (PyTorch, TensorFlow, ONNX Runtime)
  • •Hands-on performance debugging and OS/BIOS tuning, including thread/core affinity, NUMA/locality, and multi-instance execution
  • •Python/Bash automation and tooling for performance data capture, reporting, and dashboards
Experience:8-12 yearsAICPUsLinuxEPYCML inference
Education:Bachelor's
Skills:PythonBashVTuneFtraceZenDNNPyTorchTensorFlowONNX RuntimeEPYCNUMAHugePagesPerf
Languages:English
Tech Stack:PythonBashVTuneFtraceZenDNNPyTorchTensorFlowONNX RuntimeEPYCNUMAHugePages

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn