Manager I, Engineering - Coordination Systems Storage

DataDog
New York, Boston
Workplace: HybridFull timeUSD 192,000 - 240,000 annuallyFunction: Executive & General ManagementSkills: ["Collaboration","Stakeholder management","Operational excellence","Incident handling","Prioritization"]

Lead the Coordination Systems - Storage team to build and operate configuration storage and distribution systems used across Datadog services and pods. Manage a core group of five engineers in a distributed setup (majority in NYC), run ceremonies, prioritize and delegate projects, and remain hands-on with code and incident follow-ups. Drive operational excellence through reviews, root-cause analysis, and reliability-focused initiatives while partnering across teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
DataDog
DataDog
1 day ago

Manager I, Engineering - Coordination Systems Storage

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 20 hours agoStatus: Live

Job Summary

Lead the Coordination Systems - Storage team to build and operate configuration storage and distribution systems used across Datadog services and pods. Manage a core group of five engineers in a distributed setup (majority in NYC), run ceremonies, prioritize and delegate projects, and remain hands-on with code and incident follow-ups. Drive operational excellence through reviews, root-cause analysis, and reliability-focused initiatives while partnering across teams.
Location: New York, Boston
Workplace: Hybrid
Employment Type: Full time
Job Function: Executive & General Management
Seniority: Manager level

Key Responsibilities

  • •Lead the Coordination Systems - Storage team and a core group of five engineers (distributed, majority in NYC).
  • •Run engineering ceremonies and prioritize/delegate project work.
  • •Stay hands-on with code, working on isolated features and small remediations and follow-ups.
  • •Participate in operations and incidents, including root cause analysis and follow-up ownership.
  • •Promote operational excellence via operational reviews, gamedays, and proactive reliability work.

Pay and Benefits

Salary: USD 192,000 - 240,000 annually
Perks:Health InsuranceDental401kPaid LeaveEquityGym MembershipRemote Work

Key Requirements

  • •Strong distributed systems skills, including understanding failure modes and end-to-end o11y, validation testing, and simulation setup.
  • •Experience with platform teams, providing critical infrastructure to internal stakeholders.
  • •Experience responding to and owning follow-ups for significant incidents.
  • •Strong cross-team collaboration and stakeholder management skills.
  • •Strong foundation in coding with experience in performance optimizations and profiling.
Experience:Distributed systemsObservability
Skills:CollaborationStakeholder managementOperational excellenceIncident handlingPrioritization
Languages:English
Tech Stack:CodePerformance optimizationsProfiling

Company Brief

DataDog
Provides a cloud-native monitoring and observability platform that unifies metrics, traces, logs, and security signals to help engineering, operations, and security teams monitor and troubleshoot modern applications and infrastructure.
Industry: Developer Tools
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2010
Glassdoor
Glassdoor: 4.1
WebsiteLinkedIn