Student Researcher (LLM Post Training – Agent & Reinforcement Learning) - 2026 Start (PhD)
San Jose
Workplace: OnsiteFull timeFunction: Research & Scientific (R&D)Education: phdSkills: ["Research","Experimentation","Programming","Model optimization","In-depth research"]Join the Seed LLM Post Training team to research and improve post-training methods for unified multimodal large models. You’ll explore and optimize post-training technologies such as SFT, RM, and RL, and work on data construction, instruction tuning, preference alignment, and model optimization. Focus areas include advancing reasoning, coding, math, agent capabilities, and investigating future use cases through large-scale model research and systems optimization.
Loading
Loading job details...
Preparing the role view and application actions.

