Research Operations, Reinforcement Learning

Anthropic
San Francisco
Workplace: OnsiteFull timeUSD 210,000 - 260,000 annuallyFunction: Education & TrainingEducation: bachelorsSkills: ["Clear writing","Attention to detail","Strong follow-through","Strategic guidance","Candid communication"]

Embed with leadership of the reinforcement learning organization to shape priorities, agendas, and follow-through that keep researchers aligned during model launches and cross-team programs. Triage items coming to leadership, surface resourcing and technical bottlenecks early, and translate dense material into clear plans, dashboards, and org-wide communications. Identify recurring friction and act as a candid thought partner to improve how the team operates.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
1 day ago

Research Operations, Reinforcement Learning

âś“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 34 minutes agoStatus: Live

Job Summary

Embed with leadership of the reinforcement learning organization to shape priorities, agendas, and follow-through that keep researchers aligned during model launches and cross-team programs. Triage items coming to leadership, surface resourcing and technical bottlenecks early, and translate dense material into clear plans, dashboards, and org-wide communications. Identify recurring friction and act as a candid thought partner to improve how the team operates.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Education & Training
Seniority: Mid level

Key Responsibilities

  • •Shape leadership’s time by setting agendas, focusing meetings on high-leverage decisions, and cutting non-critical work.
  • •Triage inbound items: resolve what you can, route appropriately, and escalate thoughtfully.
  • •Enable fast decision-making by providing the information leaders need, communicating decisions, and tracking commitments to completion.
  • •Surface resourcing gaps, technical bottlenecks, and stalled action items before they become blockers.
  • •Scale leadership’s voice by drafting org-wide announcement and planning materials, and improve recurring operational friction.

Pay and Benefits

Salary: USD 210,000 - 260,000 annually
Perks:Paid LeaveParental Leave

Key Requirements

  • •Experience in an operations, chief of staff, program management, or comparable role supporting senior leaders in a fast-moving technical organization.
  • •Technical fluency to engage substantively with researchers, including understanding reinforcement learning training discussions and evaluation results.
  • •Ability to turn dense, jargon-heavy materials into clear summaries and plans that technical collaborators can act on.
  • •Strong follow-through and attention to detail, with an eye for dashboards and visualizations that make lessons at a glance.
  • •Ability to push back on senior stakeholders with strategic guidance, influencing without direct authority.
Experience:Reinforcement learningAI research
Education:Bachelor's
Skills:Clear writingAttention to detailStrong follow-throughStrategic guidanceCandid communication

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn