Senior SDET, Inference Platform

Cerebras
Sunnyvale, Toronto
Workplace: OnsiteFull timeFunction: QA, Test & Release EngineeringExperience: 5+ yearsSkills: ["Problem-solving","Communication","Collaboration","Debugging","Mentoring"]

Own quality and reliability of infrastructure that deploys and runs the Cerebras Inference Platform. Build and maintain test infrastructure and automation across CI/CD, Kubernetes-based deployments, and production networking components like ingress, load balancing, and service discovery. Validate in cloud and on real Cerebras clusters, collaborate with the Inference Platform development team, and debug complex distributed and orchestration issues to keep massive-scale inference fast, stable, and production-ready.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cerebras
Cerebras
2 months ago

Senior SDET, Inference Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Own quality and reliability of infrastructure that deploys and runs the Cerebras Inference Platform. Build and maintain test infrastructure and automation across CI/CD, Kubernetes-based deployments, and production networking components like ingress, load balancing, and service discovery. Validate in cloud and on real Cerebras clusters, collaborate with the Inference Platform development team, and debug complex distributed and orchestration issues to keep massive-scale inference fast, stable, and production-ready.
Location: Sunnyvale, Toronto
Workplace: Onsite
Employment Type: Full time
Job Function: QA, Test & Release Engineering
Seniority: Mid level

Key Responsibilities

  • •Design, build, and maintain test infrastructure and automation for deploying and validating the Cerebras Inference Platform.
  • •Validate the platform across environments, including cloud-managed Kubernetes and deployments running on Cerebras hardware.
  • •Test and verify deployment infrastructure including Kubernetes workloads, CI/CD pipelines, ingress and service discovery, NGINX, and load balancing.
  • •Collaborate with the Inference Platform development team to ensure new features and platform capabilities ship reliably.
  • •Investigate and debug complex issues spanning networking, orchestration, deployment, and distributed services.

Key Requirements

  • •5+ years of experience in software engineering, QA/quality engineering, systems engineering, or infrastructure development.
  • •Strong programming skills in Python and/or Go (experience with both is a plus).
  • •Experience building automation tools, testing frameworks, or internal developer tooling.
  • •Hands-on experience with CI/CD systems (e.g., Jenkins).
  • •Experience debugging complex systems, distributed services, or networked infrastructure.
Experience:5+ yearsDistributed systemsCloudCluster infrastructureInference platformsCI/CDKubernetesNetworked infrastructureML inference infrastructure
Skills:Problem-solvingCommunicationCollaborationDebuggingMentoring
Tech Stack:PythonGoJenkinsCI/CDKubernetesCloud-managed KubernetesAmazon EKSGitOpsArgoCDIngress controllersIngressService discoveryNGINXLoad balancingBazelK9sDistributed services

Company Brief

Cerebras
Designs and builds wafer-scale AI accelerators and systems for large-scale deep learning workloads, delivering specialized hardware and software to accelerate model training and inference for enterprises and research institutions.
Industry: Hardware Devices
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Sunnyvale, United States
Founded: 2016
WebsiteLinkedIn