Research Engineer, Data Infrastructure

Mistral
Paris, Zurich, Switzerland, Warsaw, London
Workplace: HybridFull timeFunction: Research & Scientific (R&D)Experience: 4+ yearsSkills: ["Problem-solving","Debugging","Scalability mindset","Reliability focus","Security mindset"]

Build and operate Mistral’s next-generation data infrastructure for frontier model training and fine-tuning. Help scale massive distributed compute and storage systems, implement multi-cluster orchestration and cloud-bursting, and transition away from legacy schedulers. Own production-grade pipelines, metadata and lineage systems, and cloud-native deployment workflows across Kubernetes and SLURM environments, including on-call rotation for critical training jobs.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Mistral
Mistral
1 month ago

Research Engineer, Data Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Build and operate Mistral’s next-generation data infrastructure for frontier model training and fine-tuning. Help scale massive distributed compute and storage systems, implement multi-cluster orchestration and cloud-bursting, and transition away from legacy schedulers. Own production-grade pipelines, metadata and lineage systems, and cloud-native deployment workflows across Kubernetes and SLURM environments, including on-call rotation for critical training jobs.
Location: Paris, Zurich, Switzerland, Warsaw, London
Workplace: Hybrid
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Mid level

Key Responsibilities

  • •Build and scale massive distributed compute and storage systems
  • •Architect and maintain multi-cluster orchestration layers to optimize workload placement across hardware and regions
  • •Design and transition to modern storage formats for large-scale fine-tuning datasets
  • •Contribute to an internal training platform enabling seamless training and fine-tuning across Kubernetes and SLURM environments
  • •Implement and manage metadata, lineage, and production deployment workflows; participate in on-call rotations for critical training jobs

Pay and Benefits

Perks:Health InsuranceParental LeaveRetirementRelocationWellness StipendMeal AllowanceCommuter Benefits

Key Requirements

  • •4+ years of experience in Data Infrastructure, MLOps, or Infrastructure Engineering
  • •Experience or strong interest supporting foundational compute and storage platforms
  • •Proficiency in Python and interest in solving “brittle data lake” problems using modern columnar storage
  • •Experience with Kubernetes-native tooling and debugging large-scale distributed systems across multi-cluster environments
  • •Comfort building scalable, reliable, secure systems in a rapid-growth AI environment
Experience:4+ yearsMLOpsData infrastructureInfrastructure engineering
Skills:Problem-solvingDebuggingScalability mindsetReliability focusSecurity mindset
Tech Stack:PythonKubernetesSLURMCloud-nativeMulti-cluster orchestrationData lakesMetadata systemsColumnar storage

Company Brief

Mistral
Develops state-of-the-art large language models and AI systems, offering models and developer tools for natural language understanding, generation, and enterprise AI integrations. Focuses on open research and production-ready model deployments.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Early Stage Startup
Funding: Seed
Headquarters: Paris, France
Founded: 2023
WebsiteLinkedIn