Research Manager

hud
San Francisco, Singapore
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Experience: 5+ yearsSkills: ["Technical judgment","Mentoring","Communication","Collaboration","Experimental judgment"]

Lead technical research that improves agent training data and evaluation quality for frontier AI models. You’ll guide research engineers through ambiguous problem spaces, define the right questions, and run rigorous experiments linking model behavior and failure modes to data, environments, and reward design. Build scalable validation methods (trajectory audits, grader checks, feedback loops) and translate findings into repeatable workflows, tools, and quality standards.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
hud
hud
2 days ago

Research Manager

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Lead technical research that improves agent training data and evaluation quality for frontier AI models. You’ll guide research engineers through ambiguous problem spaces, define the right questions, and run rigorous experiments linking model behavior and failure modes to data, environments, and reward design. Build scalable validation methods (trajectory audits, grader checks, feedback loops) and translate findings into repeatable workflows, tools, and quality standards.
Location: San Francisco, Singapore
Workplace: Remote
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Manager level

Key Responsibilities

  • •Set research direction for data quality, including reliability and usefulness of tasks, trajectories, rewards, and evals for training agents.
  • •Lead research engineers from problem definition through experiments, implementation, and conclusions; coach for technical judgment and execution.
  • •Design and review experiments connecting model behavior and failure modes to data, environment, and reward design.
  • •Develop methods for validating and improving training data at scale, including trajectory audits, grader checks, and feedback loops.
  • •Partner with research engineers, domain experts, and data vendors to turn insights into better workflows, tools, and quality standards.

Pay and Benefits

Perks:Health InsuranceDentalVision401kPaid Leave

Key Requirements

  • •Experience leading technical research projects from open question through evidence, decisions, and working results.
  • •Experience directly managing and mentoring researchers or research engineers while staying engaged in technical work.
  • •Strong understanding of machine learning and reinforcement learning, including how training objectives, data, and feedback shape behavior.
  • •Experience with agent training data, evals, benchmarks, synthetic data, or model evaluation infrastructure.
  • •Strong experimental judgment and written communication to explain methods and findings to researchers, engineers, and external partners.
Experience:5+ yearsMachine learningReinforcement learningAI agentsResearchSynthetic dataModel evaluation
Skills:Technical judgmentMentoringCommunicationCollaborationExperimental judgment
Tech Stack:Reinforcement learningMachine learningTraining objectivesTrajectory auditsGrader checksFeedback loopsChatGPTClaude CodeCursor

Eligibility

Nationality:US National
Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

hud
Hud provides a lightweight developer-focused tool that captures and shares UI components and interactive design specs from the browser to streamline design-to-development handoff and collaboration between designers and engineers.
Industry: Developer Tools
Website