Production Engineer, Network

FluidStack
San Francisco, California, New York, Austin, London, Seattle
Workplace: OnsiteFull timeUSD 175,000 - 300,000 annuallyFunction: Manufacturing & Production OperationsSkills: ["Go","Python","GNMI","GRPC","NETCONF","SONiC"]

Own network fleet health end to end, build active debugging tooling, and turn repair into an automated pipeline for a large datacenter network. You’ll define monitoring requirements, ship dashboards, and drive reliability and scalability across multiple sites in a fast-moving AI compute infrastructure company.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
FluidStack
FluidStack
3 months ago

Production Engineer, Network

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Own network fleet health end to end, build active debugging tooling, and turn repair into an automated pipeline for a large datacenter network. You’ll define monitoring requirements, ship dashboards, and drive reliability and scalability across multiple sites in a fast-moving AI compute infrastructure company.
Location: San Francisco, California, New York, Austin, London, Seattle
Workplace: Onsite
Employment Type: Full time · Permanent
Job Function: Manufacturing & Production Operations

Key Responsibilities

  • •Own network fleet health end to end. Define the realtime monitoring requirements, build the alerting lifecycle, and ship the dashboards that give every on-call engineer a true picture of network state across all sites.
  • •Build active debugging tooling. Link diagnostics, remote command execution across the fleet, and repair visualization — the tools that turn a network fault from a mystery into a solvable problem, fast.
  • •Turn repair into a pipeline, not a procedure. Build the automation that takes a network failure from detection through parts management and return to service. Ticket integration, repair lifecycle pipelines, transceiver and optics tracking — owned, not improvised.
  • •Own network qualification and validation. Build the frameworks that gate new sites and hardware into production. You define what a healthy network looks like before it carries traffic.
  • •Own end-to-end reliability, scalability, and operation of the network at-scale. Fluidstack is building one of the largest datacenter networks in the world and that can only be accomplished with aggressive automation, tooling, and incident discipline.

Pay and Benefits

Salary: USD 175,000 - 300,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceVisionDentalPensionEquity

Key Requirements

  • •Shipped production network tooling or automation and built end-to-end network reliability solutions.
  • •Proficiency in Go or Python and building automation for large-scale datacenter networks.
  • •Experience with network telemetry/tools such as link diagnostics, active debugging, and monitoring dashboards.
  • •Strong systems-thinking, problem-solving, and ability to operate with minimal supervision in a high-intensity environment.
  • •Familiarity with AI tooling and large language model integration is a plus.
Experience:DatacenterCloudAINetworking
Skills:GoPythonGNMIGRPCNETCONFSONiC
Tech Stack:GoPythonGNMIGRPCNETCONFSONiC

Company Brief

FluidStack
Builds and deploys large-scale GPU cloud infrastructure for AI labs, enterprises and governments, providing high-performance AI training and inference capacity and rapid data-center deployment services.
Industry: Cloud Computing
Company Size: Medium (51 to 250 employees)
Growth: Growth Stage Startup
Funding: Series A
Headquarters: New York, United States
Founded: 2017
Glassdoor
Glassdoor: 4.7
WebsiteLinkedInGlassdoor