Member of Technical Staff (Software Engineer)

Cerebras
Sunnyvale
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 1+ yearsEducation: mastersSkills: ["Cross-functional collaboration","Troubleshooting","Documentation","Performance optimization","Debugging"]

Develop and maintain high-performance, low-latency inference infrastructure for production machine learning services. Deploy and scale inference workloads using Kubernetes, optimize autoscaling and resource allocation, and ensure high availability through multi-region deployment and disaster recovery. Build Python APIs and scripts for real-time inference workflows, troubleshoot defects via logs/metrics/traces, and collaborate with ML and product/UX teams to validate performance, reliability, and interface requirements.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cerebras
Cerebras
4 months ago

Member of Technical Staff (Software Engineer)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live

Job Summary

Develop and maintain high-performance, low-latency inference infrastructure for production machine learning services. Deploy and scale inference workloads using Kubernetes, optimize autoscaling and resource allocation, and ensure high availability through multi-region deployment and disaster recovery. Build Python APIs and scripts for real-time inference workflows, troubleshoot defects via logs/metrics/traces, and collaborate with ML and product/UX teams to validate performance, reliability, and interface requirements.
Location: Sunnyvale
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Entry level

Key Responsibilities

  • •Implement infrastructure for high-performance, low-latency inference services.
  • •Deploy and configure Kubernetes services for scalable, reliable inference workloads.
  • •Optimize resource allocation and autoscaling policies to handle variable demand and reduce operational costs.
  • •Build Python-based scripts and APIs for data preprocessing, inference execution, and post-processing.
  • •Ensure availability and fault tolerance via multi-region deployments and disaster recovery, and triage production defects using logs, metrics, and distributed traces.

Key Requirements

  • •Master's degree (or foreign equivalent) in Computer Science or a related field.
  • •1+ year of experience as a Software Developer, Student/Intern (Software Developer), Member of Technical Staff (Software Engineer), Software Engineer, or related role.
  • •Experience gained before, during, or after graduate studies is accepted on a full-time or equivalent part-time basis.
  • •Docker and Kubernetes experience.
  • •Proficiency with at least Java or C++, plus experience with ActiveMQ and Kafka, Python or Groovy, JavaScript or TypeScript, Linux, SQL/OracleDB/Redis, and Git.
Experience:1+ yearsAIMachine learningInference infrastructureCloud inferenceDistributed systems
Education:Master's in Computer Science (or related field)
Skills:Cross-functional collaborationTroubleshootingDocumentationPerformance optimizationDebugging
Tech Stack:KubernetesDockerPythonGroovyJavaC++JavaScriptTypeScriptLinuxSQLOracleDBRedisActiveMQKafkaGitJiraAPIsJenkinsDistributed tracesMetrics

Company Brief

Cerebras
Designs and builds wafer-scale AI accelerators and systems for large-scale deep learning workloads, delivering specialized hardware and software to accelerate model training and inference for enterprises and research institutions.
Industry: Hardware Devices
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Sunnyvale, United States
Founded: 2016
WebsiteLinkedIn