Research Engineer, Model Inference & Serving - Paris
H Company
Paris, London
Workplace: HybridFull timeFunction: Software EngineeringEducation: mastersSkills: ["Collaboration","Communication","Teamwork","Problem-solving","Adaptability"]A research-oriented software engineering role focused on building and optimizing low-latency AI inference pipelines for agentic models. You will work on GPU-accelerated kernels, model compression techniques, and collaborate with research teams to push Memory/Throughput/Latency improvements in a hybrid Paris/London environment.

