Senior Site Reliability Engineer

Duolingo
New York
Workplace: OnsiteFull timeUSD 182,800 - 247,300 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Mentorship"]

Work closely with product and platform engineering teams to keep Duolingo’s distributed systems reliable, scalable, and highly operable. Identify sources of instability, support core production infrastructure, and provide system design consulting including launch reviews and root cause analysis. Maintain incident response and postmortem practices, reduce toil via tooling automation, and collaborate to release new features—becoming an authority on the services that power millions of users.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Duolingo
Duolingo
3 days ago

Senior Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 18 hours agoStatus: Live

Job Summary

Work closely with product and platform engineering teams to keep Duolingo’s distributed systems reliable, scalable, and highly operable. Identify sources of instability, support core production infrastructure, and provide system design consulting including launch reviews and root cause analysis. Maintain incident response and postmortem practices, reduce toil via tooling automation, and collaborate to release new features—becoming an authority on the services that power millions of users.
Location: New York
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Collaborate with internal teams to identify sources of instability in distributed systems and drive operational excellence.
  • •Support core infrastructure by understanding, diagnosing, and debugging production systems.
  • •Provide system design consulting, develop software platforms/frameworks, and run launch reviews with root cause analysis.
  • •Maintain and document sustainable incident response and postmortem practices.
  • •Implement changes that improve reliability, scalability, and velocity while reducing toil through tooling and automation.

Pay and Benefits

Salary: USD 182,800 - 247,300 annually
Equity and Bonus:Equity

Key Requirements

  • •5+ years of experience in site reliability engineering/DevOps for a product used by millions of users.
  • •Experience identifying and solving issues in large-scale distributed systems.
  • •Experience with Java, Kotlin, Python, or Go.
  • •Experience with containerization and container orchestration technologies (e.g., Docker, Mesos, Kubernetes, Nomad).
  • •Exceptional candidates: experience reducing maintenance toil through automation and tooling, and improving incident response processes.
Experience:Site reliability engineeringDevOpsDistributed systems
Skills:CollaborationMentorship
Languages:English
Tech Stack:JavaKotlinPythonGoDockerMesosKubernetesNomadDynamoMySQLPostgreSQL

Company Brief

Duolingo
Duolingo develops a gamified language-learning platform offering courses, practice exercises, and assessment tools for learners worldwide via web and mobile apps, using adaptive algorithms and motivational features to drive engagement and retention.
Industry: EdTech
Company Size: Enterprise (1,001+ employees)
Revenue: USD 500M to 1B
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Pittsburgh, United States
Founded: 2011
Glassdoor
Glassdoor: 4.2
WebsiteLinkedInGlassdoor