Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

MongoDB
Montreal, Toronto, New York
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 6+ yearsSkills: ["Problem-solving","Collaboration","Communication"]

Senior Site Reliability Engineer on MongoDB’s Storage Layer Services (SLS) team, helping re-architect a cloud storage layer for Atlas. You’ll define SLOs, shape capacity plans, and ensure reliability, durability, and safety of multi-tenant storage. Join a small, senior team in a multi-year roadmap, with flexibility to work from Montreal, Toronto, or remotely in Canada in Eastern/Central time zones while building scalable infrastructure.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
MongoDB
MongoDB
5 months ago

Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 hours agoStatus: Live

Job Summary

Senior Site Reliability Engineer on MongoDB’s Storage Layer Services (SLS) team, helping re-architect a cloud storage layer for Atlas. You’ll define SLOs, shape capacity plans, and ensure reliability, durability, and safety of multi-tenant storage. Join a small, senior team in a multi-year roadmap, with flexibility to work from Montreal, Toronto, or remotely in Canada in Eastern/Central time zones while building scalable infrastructure.
Location: Montreal, Toronto, New York
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Sr. Manager level

Key Responsibilities

  • •Work on multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs
  • •Build for reliability, making services and infrastructure available, resilient, fault-tolerant, and self-healing
  • •Identify and configure key metrics to detect incidents and quantify service health, availability, and performance
  • •Participate in a 24/7 on-call rotation to resolve issues involving the storage infrastructure
  • •Become an expert in infrastructure performance, helping optimize from the application level to the kernel

Key Requirements

  • •Have 6+ years of experience working on software development and operating distributed systems
  • •Proficiency in Python, Go, or a similar language
  • •Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs.
  • •Possess a customer-focused mindset
  • •Value efficiency in processes and operations
Experience:6+ yearsCloudDistributed systemsStorageDatabases
Skills:Problem-solvingCollaborationCommunication
Languages:English
Tech Stack:PythonGoKubernetesAWSGCPAzureLinuxTCP/IPDNSTLS

Company Brief

MongoDB
Develops MongoDB, a leading general-purpose, document-based database platform that enables developers and enterprises to build scalable, high-performance applications with flexible data models and cloud-native capabilities.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2007
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn