AI Engineer – Model Optimization & Acceleration

AMD
Bengaluru
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningSkills: ["Problem-solving","Communication","Collaboration"]

Lead the optimization and deployment of ML models across CPU, GPU, and NPU platforms, delivering production-ready AI systems for robotics, healthcare, and automotive applications. You will optimize models, port across frameworks, and improve latency and memory usage while targeting diverse hardware accelerators. Collaborate with cross-functional teams to ship features used across AMD's AI workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AMD
AMD
3 months ago

AI Engineer – Model Optimization & Acceleration

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live
Reposted: similar role first listed 3 months ago

Job Summary

Lead the optimization and deployment of ML models across CPU, GPU, and NPU platforms, delivering production-ready AI systems for robotics, healthcare, and automotive applications. You will optimize models, port across frameworks, and improve latency and memory usage while targeting diverse hardware accelerators. Collaborate with cross-functional teams to ship features used across AMD's AI workloads.
Location: Bengaluru
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Optimize diverse models: generative (LLMs, diffusion), vision (classification, detection, segmentation), multi-modal, and speech
  • •Port models across frameworks (e.g., PyTorch → ONNX → runtimes)
  • •Deploy on hardware accelerators (GPU/NPU) and optimize performance
  • •Improve inference latency, throughput, and memory (batching, caching, parallelism, fusion)
  • •Apply quantization and model compression (FP32 → lower precision)
  • •Profile and debug system and model performance

Key Requirements

  • •Strong in PyTorch (or similar), ONNX (or equivalent)
  • •Proficient in Python and C++
  • •Experience with GPU/hardware acceleration (CUDA/ROCm or similar)
  • •Solid understanding of deep learning models (transformers, CNNs)
  • •Knowledge of optimization, quantization, and performance tuning
Experience:AIMLRoboticsHealthcareAutomotive
Skills:Problem-solvingCommunicationCollaboration
Languages:English
Tech Stack:PyTorchONNXPythonC++CUDAGPU

Company Brief

AMD
Designs and produces semiconductor products including CPUs, GPUs, and adaptive SoCs for consumer, enterprise, and embedded markets, competing across PCs, data centers, and gaming industries.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1969
Glassdoor
Glassdoor: 3.9
WebsiteLinkedIn