Applied Scientist

Adobe Systems
San Jose, Seattle
Workplace: OnsiteFull timeUSD 120,700 - 238,600 annuallyFunction: Research & Scientific (R&D)Education: phdSkills: ["Communication","Collaboration"]

Build and improve mid-training methods for Adobe multimodal generative models used for image, video, and audio editing. Own components of the training stack, design and run experiments to close quality gaps, and develop large-scale captioning pipelines to support VLM finetuning. Collaborate with research, data, evaluation, and infrastructure teams to deliver scalable workflows for data curation and distributed training, impacting creative workflows at millions of users.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Adobe Systems
Adobe Systems
4 months ago

Applied Scientist

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Build and improve mid-training methods for Adobe multimodal generative models used for image, video, and audio editing. Own components of the training stack, design and run experiments to close quality gaps, and develop large-scale captioning pipelines to support VLM finetuning. Collaborate with research, data, evaluation, and infrastructure teams to deliver scalable workflows for data curation and distributed training, impacting creative workflows at millions of users.
Location: San Jose, Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Mid level

Key Responsibilities

  • •Design, implement, and evaluate mid-training approaches that improve editing capabilities for Adobe multimodal generative models across image, video, and audio.
  • •Own defined components within the mid-training stack (e.g., image-to-image editing or instruction-based editing) and run experiments to test hypotheses and identify quality gaps.
  • •Build and maintain large-scale captioning pipelines and support VLM finetuning to improve multimodal understanding across visual and auditory domains.
  • •Assist in building scalable workflows for data curation, quality improvements, and distributed training using research insights from diffusion models and large-scale training.
  • •Collaborate with cross-functional teams across research, data, evaluation, and training to document and share experiment and dataset learnings.

Pay and Benefits

Salary: USD 120,700 - 238,600 annually

Key Requirements

  • •Possess a Master’s or Ph.D. degree in Computer Science, Machine Learning, or a related field.
  • •Solid understanding of modern generative architectures such as diffusion models, including conditional generation or editing methods.
  • •Experience implementing machine learning models using deep learning frameworks such as PyTorch, including large-scale or distributed training workflows.
  • •Experience or research background in VLM finetuning for image, video, and audio understanding and adapting pretrained vision-language models for multimodal tasks.
  • •Familiarity with large-scale captioning, including designing or applying automated captioning pipelines.
Experience:Generative AIMultimodalVLM finetuningLarge-scale trainingDiffusion models
Education:PhD / Doctorate in Computer Science, Machine Learning
Skills:CommunicationCollaboration
Tech Stack:PythonPyTorchDiffusion modelsVision-language modelsVLM finetuning

Company Brief

Adobe Systems
Provides creative, marketing, and document management software and cloud services, including Photoshop, Illustrator, Acrobat, and the Adobe Experience Cloud, serving creative professionals, enterprises, and governments worldwide.
Industry: SaaS
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: San Jose, United States
Founded: 1982
Glassdoor
Glassdoor: 4.0
WebsiteLinkedInGlassdoor