Senior Site Reliability Engineer

MongoDB
Gurugram
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureExperience: 6+ yearsSkills: ["Technical communication","Customer focus","Collaboration","Process efficiency","Automation mindset"]

Build and run the operational foundations for a new platform enabling customers to build AI applications with MongoDB. Lead deployment at scale by improving performance, scalability, and reliability of distributed infrastructure. Own and evolve Kubernetes fleet operations, networking, observability and alerting, and tenant isolation, while identifying key metrics and participating in a 24/7 on-call rotation. Mentor early-career SREs as the team grows.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
MongoDB
MongoDB
1 month ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Build and run the operational foundations for a new platform enabling customers to build AI applications with MongoDB. Lead deployment at scale by improving performance, scalability, and reliability of distributed infrastructure. Own and evolve Kubernetes fleet operations, networking, observability and alerting, and tenant isolation, while identifying key metrics and participating in a 24/7 on-call rotation. Mentor early-career SREs as the team grows.
Location: Gurugram
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Provide technical direction and help shape architecture for a new platform enabling AI application development with MongoDB.
  • •Operate and improve multi-tenant Kubernetes infrastructure that runs customer workloads.
  • •Build for reliability with services and infrastructure that are resilient, fault-tolerant, and self-healing.
  • •Configure key metrics for incident detection and to quantify service health, availability, and performance.
  • •Participate in a 24/7 on-call rotation and mentor early-career SREs as the team grows.

Pay and Benefits

Perks:Parental Leave

Key Requirements

  • •Strong background in software development and operating distributed systems.
  • •6+ years building and operating distributed systems using Python, Go, or a similar programming language.
  • •Operate Kubernetes in production and debug below abstraction layers, including scheduling, cluster networking, and node-level issues.
  • •Expertise in cloud infrastructure platforms including AWS, Google Cloud Platform (GCP), or Azure.
  • •Strong understanding of Linux internals and networking concepts such as TCP/IP, DNS, TLS, and routing.
Experience:6+ years
Skills:Technical communicationCustomer focusCollaborationProcess efficiencyAutomation mindset
Languages:English
Tech Stack:KubernetesPythonGoAWSGoogle Cloud Platform (GCP)AzureLinuxTCP/IPDNSTLSIstioCiliumService meshEdge load balancingVirtualizationWorkload isolationMulti-cloud

Company Brief

MongoDB
Develops MongoDB, a leading general-purpose, document-based database platform that enables developers and enterprises to build scalable, high-performance applications with flexible data models and cloud-native capabilities.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2007
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn