Senior Machine Learning Engineer, Services/MLOps

Adobe Systems
San Jose, Seattle, San Francisco
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningExperience: 5+ yearsSkills: ["Python","PyTorch","TensorFlow","Docker","ML Ops","CUDA","GPU","CUDA kernels","Distributed systems"]

Lead the design and optimization of scalable ML infrastructure for large foundation models. You’ll build training pipelines across thousands of GPUs, optimize GPU utilization and latency, and ensure robust, end-to-end ML workflows in a fast-paced, startup-like environment at Adobe.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Adobe Systems
Adobe Systems
4 months ago

Senior Machine Learning Engineer, Services/MLOps

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 hours agoStatus: Live

Job Summary

Lead the design and optimization of scalable ML infrastructure for large foundation models. You’ll build training pipelines across thousands of GPUs, optimize GPU utilization and latency, and ensure robust, end-to-end ML workflows in a fast-paced, startup-like environment at Adobe.
Location: San Jose, Seattle, San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Sr. Manager level

Key Responsibilities

  • •Build and optimize infrastructures that power large foundation model training on thousands of GPUs.
  • •Profile GPU utilization, trace inference and training runs, and craft strategies to optimize ML model latency.
  • •Architect and optimize end-to-end ML pipelines to be scalable, efficient, and robust.
  • •Dive deep into data to recommend the right models, evaluation metrics, and governance approaches.
  • •Engage in architecture, design, deployment, and optimizations of ML models and systems throughout the product lifecycle.

Key Requirements

  • •5+ years ML Engineering experience, specializing in generative AI like LLMs.
  • •Strong Python and deep learning engineering skills, with experience in training and inference with PyTorch or TensorFlow.
  • •Familiarity with distillation, transformers, and diffusion models; experience with generative image and video is a plus.
  • •Knowledge of deployment technologies such as Docker, ML Ops, and ML services; experience with Azure and AWS is a plus.
  • •Graduate, PhD, or postgraduate degree in Computer Science, Computer Engineering, or related field—or equivalent experience.
Experience:5+ yearsGenerative AIFoundation modelsLLMs
Skills:PythonPyTorchTensorFlowDockerML OpsCUDAGPUCUDA kernelsDistributed systems
Languages:English
Tech Stack:PythonPyTorchTensorFlowDockerML OpsAzureAWSCUDACUDA kernels

Company Brief

Adobe Systems
Provides creative, marketing, and document management software and cloud services, including Photoshop, Illustrator, Acrobat, and the Adobe Experience Cloud, serving creative professionals, enterprises, and governments worldwide.
Industry: SaaS
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 1982
Glassdoor
Glassdoor: 4.0
WebsiteLinkedInGlassdoor