Principal Research Engineer, Model Training & Post-Training
Inflection AI
Palo Alto
Workplace: OnsiteFull timeUSD 400,000 - 550,000 annuallyFunction: Education & TrainingEducation: phdSkills: ["Transformer models","Distributed training","SFT","RLHF","DPO","GRPO","RLAIF","Reward modeling","Tool-use fine-tuning","Data quality","Evaluation design"]Lead the end-to-end model-improvement loop for large-scale foundation models, from data curation and training through evaluation, post-training, release criteria, and production feedback. Own architecture, distillation, and alignment methods; drive large-scale distributed training (1,000+ GPUs) and data strategies, ensuring high-quality, enterprise-ready models with strong reliability and cost-performance balance.

