Research Engineer / Scientist, Post-training & Reinforcement Learning - London

H Company
London
Workplace: HybridFull timeFunction: Research & Scientific (R&D)Education: mastersSkills: ["Communication","Collaboration","Problem-solving"]

Join the Models team to develop and train advanced LLMs and VLMs, focusing on multimodal architectures, distributed training, and alignment. Collaborate across teams to deploy agentic AI systems, publish research, and stay at the forefront of LLM/VLM advancements.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
H Company
H Company
5 months ago

Research Engineer / Scientist, Post-training & Reinforcement Learning - London

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Join the Models team to develop and train advanced LLMs and VLMs, focusing on multimodal architectures, distributed training, and alignment. Collaborate across teams to deploy agentic AI systems, publish research, and stay at the forefront of LLM/VLM advancements.
Location: London
Workplace: Hybrid
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Develop and train advanced LLMs and VLMs, including multimodal architectures
  • •Research and implement training methods for enhanced capabilities like instruction following and tool use
  • •Design and optimize data pipelines and training systems for large-scale distributed training
  • •Collaborate with cross-functional teams to integrate models into agentic AI systems
  • •Evaluate model performance and communicate findings to stakeholders
Travel: Medium travel

Key Requirements

  • •Strong programming skills (Python, Git)
  • •Expertise in deep learning frameworks (PyTorch, JAX, TensorFlow)
  • •Experience with large-scale distributed training of LLMs and VLMs
  • •Hands-on experience with LLM training, alignment, and reinforcement learning
  • •Knowledge of multimodal architectures and applications
Experience:Artificial intelligenceMachine learningDeep learning
Education:Master's
Skills:CommunicationCollaborationProblem-solving
Tech Stack:PythonGitPyTorchJAXTensorFlowLLMsVLMsDistributed trainingReinforcement learning

Company Brief

H Company
Develops agentic AI foundation models and deployable AI agents (e.g., Surfer H, Runner H) to automate web and enterprise tasks, improving productivity for large organisations through visual-language and planning capabilities.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Early Stage Startup
Funding: Seed
Headquarters: Paris, France
Founded: 2023
Glassdoor
Glassdoor: 2.5
WebsiteLinkedInGlassdoor