Senior Staff AI Accelerator Performance Architect

Cerebras
Sunnyvale
Workplace: OnsiteFull timeUSD 175,000 - 275,000 annuallyFunction: Hardware, Embedded & Systems EngineeringExperience: 7+ yearsEducation: mastersSkills: ["Communication","Quantitative analysis","Problem-solving"]

Guide the evolution of next-generation AI accelerator and system architectures by building and validating performance models across workloads. Analyze end-to-end inference and training execution to pinpoint bottlenecks in time, bandwidth, compute, capacity, and energy efficiency. Evaluate proposed architectural features, partner with architecture/compiler/kernel/runtime teams, and translate modeling results into trusted workload insights, ROI, and actionable recommendations.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cerebras
Cerebras
11 hours ago

Senior Staff AI Accelerator Performance Architect

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Guide the evolution of next-generation AI accelerator and system architectures by building and validating performance models across workloads. Analyze end-to-end inference and training execution to pinpoint bottlenecks in time, bandwidth, compute, capacity, and energy efficiency. Evaluate proposed architectural features, partner with architecture/compiler/kernel/runtime teams, and translate modeling results into trusted workload insights, ROI, and actionable recommendations.
Location: Sunnyvale
Workplace: Onsite
Employment Type: Full time
Job Function: Hardware, Embedded & Systems Engineering

Key Responsibilities

  • •Own and evolve performance models and modeling methodologies for next-generation accelerator and system architectures.
  • •Build analytical, simulation-based, or trace-driven models across workloads, architectural features, and product generations.
  • •Analyze AI workloads from kernels to end-to-end inference and training to determine where resources are spent.
  • •Identify hardware and software bottlenecks and quantify opportunities to improve latency, throughput, utilization, and energy efficiency.
  • •Evaluate proposed architectural features and recommend improvements based on representative workload ROI.

Pay and Benefits

Salary: USD 175,000 - 275,000 annually
Equity and Bonus:Equity

Key Requirements

  • •7+ years of experience in performance analysis, performance modeling, or architecture exploration for CPUs, GPUs, AI accelerators, or HPC systems.
  • •Strong understanding of hardware architecture developed through hardware, compiler, kernel, runtime, or system-performance work.
  • •Experience developing analytical, simulation-based, or trace-driven performance models using Python, C++, or similar environments.
  • •Solid understanding of processor architecture, memory systems, interconnects, parallel execution, and hardware resource constraints.
  • •MS or PhD in Electrical Engineering, Computer Engineering, Computer Science, or equivalent practical experience.
Experience:7+ yearsAI acceleratorsHigh-performance computingHPCTransformersAI inferenceAI training
Education:Master's
Skills:CommunicationQuantitative analysisProblem-solving
Tech Stack:PythonC++RTL simulationEmulationFPGASilicon measurementsTransformer inferenceGEMMGEMVMixture-of-expertsQuantizationCollective communication

Company Brief

Cerebras
Designs and builds wafer-scale AI accelerators and systems for large-scale deep learning workloads, delivering specialized hardware and software to accelerate model training and inference for enterprises and research institutions.
Industry: Hardware Devices
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: Sunnyvale, United States
Founded: 2016
WebsiteLinkedIn