AI Research Scientist, Reinforcement Learning (LLM) and Post-Training
Santa Clara
Workplace: HybridFull timeUSD 178,500 - 306,000 annuallyFunction: Data Science & Machine LearningEducation: phdSkills: ["Research","Analysis","Collaboration","Technical leadership","Evaluation"]Develop and advance reinforcement learning methods for post-training large language models and code models used in engineering-adjacent tasks. Invent and analyze RL algorithms such as policy optimization, preference-based methods, exploration, credit assignment, and reward modeling, then run rigorous empirical studies. Design reward models, training recipes, and curricula, characterize failure modes, and collaborate with RL infrastructure to scale training and improve measurable task success while maintaining stability and safety.
Loading
Loading job details...
Preparing the role view and application actions.

