Member of Technical Staff, AI Engineering

Micron Technology
Boise, Idaho
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningExperience: 10+ yearsEducation: bachelorsSkills: ["Communication","Mentoring","Collaboration","Problem-solving"]

Build and optimize large-scale AI training and fine-tuning workloads for GPU-based, multi-node systems. Architect custom SFT/RLHF pipelines, improve throughput and memory efficiency with distributed training and mixed precision, and develop performance-focused tooling and regression tests. Work across AI, data science, and hardware teams to analyze LLM and rendering workloads, write high-performance CUDA/HIP kernels, and collaborate on next-generation GPU features—driving performance, cost, and speed across Micron’s AI-powered manufacturing stack.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Micron Technology
Micron Technology
1 month ago

Member of Technical Staff, AI Engineering

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 21 hours agoStatus: Live

Job Summary

Build and optimize large-scale AI training and fine-tuning workloads for GPU-based, multi-node systems. Architect custom SFT/RLHF pipelines, improve throughput and memory efficiency with distributed training and mixed precision, and develop performance-focused tooling and regression tests. Work across AI, data science, and hardware teams to analyze LLM and rendering workloads, write high-performance CUDA/HIP kernels, and collaborate on next-generation GPU features—driving performance, cost, and speed across Micron’s AI-powered manufacturing stack.
Location: Boise, Idaho
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Architect and deliver large-scale custom model training and fine-tuning jobs (SFT, RLHF) on multi-node, multi-GPU clusters.
  • •Optimize training throughput and memory efficiency using distributed training strategies (FSDP, DeepSpeed, Megatron-LM) and mixed-precision techniques (FP16/BF16).
  • •Design and develop autonomous AI agents for multi-step reasoning, planning, and tool execution to automate manufacturing workflows.
  • •Analyze and profile complex workloads (e.g., LLM training, rendering pipelines) to identify bottlenecks in compute, memory bandwidth, and latency.
  • •Write and optimize high-performance kernels using CUDA/HIP/custom assembly and build performance regression testing suites; mentor engineers in parallel optimization techniques.

Pay and Benefits

Perks:Health InsuranceDentalVisionPaid Time-offPaid HolidaysParental Leave

Key Requirements

  • •10+ years of GPU architecture expertise, including memory hierarchy, tensor cores, NVLink, and GPU resource management across cloud and on-prem environments.
  • •5+ years in performance optimization, parallel computing, and low-level systems development using C++ and GPGPU frameworks (CUDA preferred; HIP/OpenCL/Metal acceptable).
  • •Hands-on experience building scalable ML systems, including distributed training (DDP, FSDP), model parallelism, and end-to-end automation of training, testing, and deployment workflows.
  • •Deep proficiency in LLMs, including prompt engineering, tool/function calling, fine-tuning with PEFT (LoRA, QLoRA), and inference optimization using vLLM and TensorRT-LLM.
  • •Experience with GenAI applications and AI agents using frameworks such as LangChain, LangGraph, LlamaIndex, and AutoGen, plus ML frameworks (PyTorch required).
Experience:10+ yearsGPU architectureDistributed trainingLLMsGenAIAI agentsHPC
Education:Bachelor's in Computer Science, Statistics, or a related field
Skills:CommunicationMentoringCollaborationProblem-solving
Tech Stack:SFTRLHFFSDPDeepSpeedMegatron-LMFP16BF16CUDAHIPPTXSASSC++GPGPUDistributed trainingDDPModel parallelismPythonJavaCI/CDGit

Company Brief

Micron Technology
Designs and manufactures semiconductor memory and storage solutions, including DRAM, NAND, and NOR flash, for computing, networking, mobile, and automotive markets worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Boise, United States
Founded: 1978
Glassdoor
Glassdoor: 3.7
WebsiteLinkedIn