Site Reliability Engineer

Anduril
United States
Workplace: OnsiteFull timeUSD 112,000 - 149,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsSkills: ["Communication","Troubleshooting","Problem-solving","Documentation","Calm under pressure"]

Own the health and uptime of deployed imaging systems as the frontline SRE for field incidents. Triage and diagnose issues across the stack, run escalations from support channels, and turn recurring problems into runbooks, diagnostics, and self-service tooling. Partner with Mission Software Engineers by reproducing true defects and feeding reliability learnings back into the product to improve observability, upgrades, and failure handling.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anduril
Anduril
2 days ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Own the health and uptime of deployed imaging systems as the frontline SRE for field incidents. Triage and diagnose issues across the stack, run escalations from support channels, and turn recurring problems into runbooks, diagnostics, and self-service tooling. Partner with Mission Software Engineers by reproducing true defects and feeding reliability learnings back into the product to improve observability, upgrades, and failure handling.
Location: United States
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Own the health and uptime of deployed imaging systems by triaging, diagnosing, and driving resolution when issues arise.
  • •Run point on escalations as first and second line of response for issues coming through support channels and customer-support pipeline.
  • •Turn recurring issues into runbooks, diagnostics, and self-service tooling to reduce repeat problems and support load.
  • •Hold the boundary with engineering by reproducing genuine software defects, documenting them, and handing them off to Mission Software Engineers.
  • •Feed reliability learnings back into the product to improve observability, safer upgrades, and more graceful failure.
Travel: Medium travel

Pay and Benefits

Salary: USD 112,000 - 149,000 annually
Equity and Bonus:Equity

Key Requirements

  • •3+ years in SRE, DevOps, field/systems engineering, or production support of deployed hardware/software systems, with real ownership after systems ship.
  • •Strong Linux fundamentals and comfort troubleshooting real networking issues (IP, routing, VPNs, connectivity in constrained/field environments).
  • •Ability to diagnose and resolve issues across system boundaries (networking, services, hardware interaction) without full visibility into every component.
  • •Comfort owning a structured on-call rotation, including scheduled after-hours and weekend coverage.
  • •Strong written and verbal communication skills for remote troubleshooting, including documenting outcomes.
Experience:3+ yearsSREDevOpsField systemsProduction supportDeployed hardware/software
Skills:CommunicationTroubleshootingProblem-solvingDocumentationCalm under pressure
Languages:English
Tech Stack:LinuxIPRoutingVPNPythonBashNixNixOSSystemdObservabilityPagerDuty

Eligibility

Security Clearance:Secret

Company Brief

Anduril
Designs and builds advanced defense systems combining autonomous aircraft, sensors, and AI-driven software for military and national security applications, focused on modernizing battlefield capabilities and distributed sensing.
Industry: Defense Technology
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: Costa Mesa, United States
Founded: 2017
WebsiteLinkedIn