Research Engineer, AI Safety & Alignment

Character.ai
Redwood City
Workplace: OnsiteFull timeUSD 225,000 - 400,000 annuallyFunction: Research & Scientific (R&D)Education: phdSkills: ["Problem-solving","Collaboration","Communication"]

Research Engineer focused on AI safety and alignment, bridging theoretical safety research with production-ready code. Develop evaluation metrics for model safety, advance alignment techniques, and collaborate with engineering and product teams to deploy safe, scalable solutions while contributing to the academic community through publications.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Character.ai
Character.ai
11 months ago

Research Engineer, AI Safety & Alignment

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Research Engineer focused on AI safety and alignment, bridging theoretical safety research with production-ready code. Develop evaluation metrics for model safety, advance alignment techniques, and collaborate with engineering and product teams to deploy safe, scalable solutions while contributing to the academic community through publications.
Location: Redwood City
Workplace: Onsite
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Develop and implement novel evaluation methodologies and metrics to assess the safety and alignment of large language models.
  • •Research and develop cutting-edge techniques for model alignment, value learning, and interpretability.
  • •Conduct adversarial testing to proactively uncover potential vulnerabilities and failure modes in our models.
  • •Analyze and mitigate biases, toxicity, and other harmful behaviors in large language models through techniques like reinforcement learning from human feedback (RLHF) and fine-tuning.
  • •Collaborate with engineering and product teams to translate safety research into practical, scalable solutions and best practices.

Pay and Benefits

Salary: USD 225,000 - 400,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Hold a PhD (or equivalent experience) in Computer Science, Machine Learning, or a related discipline.
  • •Write clear production-facing and training code.
  • •Experience working with GPUs (training, serving, debugging).
  • •Experience with data pipelines and data infrastructure.
  • •Strong understanding of modern ML techniques, particularly transformers and reinforcement learning, with a focus on safety implications.
Experience:AI safetyMachine learningAI alignment
Education:PhD / Doctorate
Skills:Problem-solvingCollaborationCommunication
Tech Stack:TransformersReinforcement learningRLHFGPUsDistributed trainingKubernetesDockerCloud

Company Brief

Character.ai
Builds a conversational AI platform where users create and interact with lifelike AI characters for entertainment, learning, and productivity, backed by advanced large‑language‑model research and consumer-facing products.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series A
Headquarters: Menlo Park, United States
Founded: 2021
Glassdoor
Glassdoor: 3.6
WebsiteLinkedInGlassdoor