Member of Technical Staff, Reinforcement Learning

Outerport
San Francisco, Tokyo
Workplace: OnsiteFull timeUSD 100,000 - 200,000 annuallyFunction: Education & TrainingSkills: ["Communication","Problem-solving","Collaboration"]

Join a founding team building reinforcement learning–based LLM training pipelines for materials science and engineering. You’ll train LLMs, integrate simulation tools and hardware to collect data, and evaluate and deploy models while collaborating with vendors and researchers. This role blends ML research with hands-on engineering, data sourcing, and an interest in manufacturing and semiconductors.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Outerport
Outerport
4 months ago

Member of Technical Staff, Reinforcement Learning

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 50 minutes agoStatus: Live

Job Summary

Join a founding team building reinforcement learning–based LLM training pipelines for materials science and engineering. You’ll train LLMs, integrate simulation tools and hardware to collect data, and evaluate and deploy models while collaborating with vendors and researchers. This role blends ML research with hands-on engineering, data sourcing, and an interest in manufacturing and semiconductors.
Location: San Francisco, Tokyo
Workplace: Onsite
Employment Type: Full time
Job Function: Education & Training

Key Responsibilities

  • •Train reinforcement learning–based LLMs to solve tasks in the domain of materials science, chemical engineering, and engineering science
  • •Integrate simulation tools and real hardware to collect data
  • •Source data from vendors and evaluate data quality
  • •Evaluate, test, and deploy models
  • •Collaborate with the founding team and cross-functional partners

Pay and Benefits

Salary: USD 100,000 - 200,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Experience building reinforcement learning-based LLM training pipelines (Python / PyTorch)
  • •Industry ML experience on hard problems
  • •Publication record in venues like ICLR, NeurIPS, ICML
  • •Strong communication skills
  • •Experience with engineering simulation tools (CAE, EDA)
Experience:Machine learningReinforcement learningLLMsMaterials science
Skills:CommunicationProblem-solvingCollaboration
Tech Stack:PythonPyTorchTorch/PyTorchCAEEDA tools

Eligibility

Visa:US citizen/visa only
Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Outerport
Provides digital freight forwarding and international shipping services, offering end-to-end logistics coordination, carrier booking, customs support, and shipment tracking to simplify cross-border supply chain management for businesses.
Industry: Shipping & Freight
Website