Software Engineer, GenAI Platform

Deliveroo
London
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 3+ yearsEducation: bachelorsSkills: ["Adaptability","Collaboration","Problem-solving","Operational excellence"]

Build production infrastructure for a centralized GenAI open-weights model platform, focusing on real-time GPU serving, high-throughput batch inference, and fine-tuning workflows (SFT/DPO/LoRA). Design scalable, high-performance systems for autoscaling GPUs, model-serving reliability, observability, and cost controls. Partner with ML, product, data science, and platform teams across Deliveroo, DoorDash, and Wolt to turn GenAI prototypes into durable platform primitives.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Deliveroo
Deliveroo
1 month ago

Software Engineer, GenAI Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 hours agoStatus: Live

Job Summary

Build production infrastructure for a centralized GenAI open-weights model platform, focusing on real-time GPU serving, high-throughput batch inference, and fine-tuning workflows (SFT/DPO/LoRA). Design scalable, high-performance systems for autoscaling GPUs, model-serving reliability, observability, and cost controls. Partner with ML, product, data science, and platform teams across Deliveroo, DoorDash, and Wolt to turn GenAI prototypes into durable platform primitives.
Location: London
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Build infrastructure to move GenAI ideas from prototype to production and increase business impact velocity.
  • •Develop the open-weights serving stack, including real-time GPU endpoints, high-throughput batch inference, and fine-tuning workflows (SFT/DPO/LoRA).
  • •Design scalable, high-performance systems for model serving, GPU autoscaling/utilization, batch pipelines, and backend services.
  • •Improve cost and latency of GPU inference and enable reliable model selection with fallback, observability, and cost controls.
  • •Collaborate with ML, product, data science, and platform teams to deliver reusable platform primitives and evolve the centralized GenAI platform (including post-training and agentic techniques).

Key Requirements

  • •BSc, MSc, or PhD in Computer Science (or equivalent).
  • •3+ years of industry experience in software engineering.
  • •Strong backend fundamentals, especially in Python and distributed systems.
  • •Experience building production services, APIs, data pipelines, or ML infrastructure at scale.
  • •Hands-on experience with LLM inference and/or fine-tuning of open-weight models in production (serving and/or SFT/DPO/LoRA).
Experience:3+ yearsGenerative AILLM inferenceModel fine-tuningGPU infrastructureDistributed systems
Education:Bachelor's in Computer Science
Skills:AdaptabilityCollaborationProblem-solvingOperational excellence
Tech Stack:PythonDistributed systemsAPIsData pipelinesLLMVLMOpen-weight modelsGPU servingBatch inferenceFine-tuningSFTDPOLoRALLM GatewayAgent GatewayEvals infrastructureGuardrailsCost attributionGPU autoscalingGPU utilization

Company Brief

Deliveroo
Deliveroo is a multinational online food and grocery delivery marketplace connecting consumers, restaurants, retailers and riders via its app and platform, operating across multiple countries with fast local delivery and logistics services.
Industry: Online Marketplaces
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Unicorn (USD 1B+)
Funding: IPO / Publicly Listed
Headquarters: London, United Kingdom
Founded: 2013
Glassdoor
Glassdoor: 3.4
WebsiteLinkedInGlassdoor