Research Engineer, Cybersecurity RL (Reinforcement Learning)

Anthropic
Zurich
Workplace: OnsiteFull timeFunction: CybersecurityEducation: bachelorsSkills: ["Collaboration","Communication"]

Build defensive cybersecurity capabilities by combining ML research with production-ready engineering. Design and implement reinforcement learning environments for incident response, security analysis, and vulnerability remediation. Conduct experiments and evaluations, deliver work into training runs, and develop agentic integrations that let AI systems autonomously investigate analytical findings. Collaborate closely with safety teams and security specialists, with possible on-call participation.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
2 days ago

Research Engineer, Cybersecurity RL (Reinforcement Learning)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Build defensive cybersecurity capabilities by combining ML research with production-ready engineering. Design and implement reinforcement learning environments for incident response, security analysis, and vulnerability remediation. Conduct experiments and evaluations, deliver work into training runs, and develop agentic integrations that let AI systems autonomously investigate analytical findings. Collaborate closely with safety teams and security specialists, with possible on-call participation.
Location: Zurich
Workplace: Onsite
Employment Type: Full time
Job Function: Cybersecurity

Key Responsibilities

  • •Partner with researchers and safety teams to understand analytical needs and build solutions.
  • •Develop agentic integrations so AI systems can autonomously investigate and act on analytical findings.
  • •Contribute to strategic team direction, including what to build, what to partner on, and where to invest.
  • •Design and implement RL environments for defensive cybersecurity use cases.
  • •Participate in an on-call rotation if required.

Pay and Benefits

Perks:Parental LeaveEquity

Key Requirements

  • •Experience with machine learning.
  • •Experience with cybersecurity research.
  • •Strong software engineering skills.
  • •Ability to balance research exploration with engineering implementation.
  • •Familiarity with RL techniques and environments (and LLM training methodologies).
Experience:CybersecurityAI safetyReinforcement learning
Education:Bachelor's
Skills:CollaborationCommunication
Languages:English
Tech Stack:Reinforcement learningRLLLM training

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn