Site Reliability Engineer
Mistral
New York
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 7+ yearsEducation: mastersSkills: ["Problem-solving","Communication","Collaboration","Self-motivated","Knowledge sharing"]Shape the reliability, scalability, and performance of production systems and customer-facing applications. Balance day-to-day SRE operations with long-term engineering improvements to reduce operational toil. Design and maintain fault-tolerant infrastructure for web services and ML workloads, build monitoring/incident response, and develop CI/CD and orchestration workflows using modern tools. Collaborate with software and AI/ML research teams to enable safe, reproducible experiments and cloud-agnostic platform capabilities.

