Software Engineer, AI accelerator Runtime

OpenAI
San Francisco
Workplace: HybridFull timeUSD 266,000 - 445,000 annuallyFunction: Software EngineeringSkills: ["Debugging","Collaboration","Quantitative reasoning","Problem-solving"]

Build OpenAI’s low-level device runtime for custom AI accelerator hardware, turning compiled programs into efficient execution. You’ll implement kernel-launch scheduling, command submission, dependency tracking, and robust memory management across concurrent workloads. Partner with compiler, kernel, architecture, and silicon teams to define interfaces and validate behavior using event-based, cycle-accurate simulators, then diagnose correctness and performance issues. Add tests, tracing, profiling, and observability to ensure production-ready runtime capabilities.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
2 days ago

Software Engineer, AI accelerator Runtime

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 hours agoStatus: Live

Job Summary

Build OpenAI’s low-level device runtime for custom AI accelerator hardware, turning compiled programs into efficient execution. You’ll implement kernel-launch scheduling, command submission, dependency tracking, and robust memory management across concurrent workloads. Partner with compiler, kernel, architecture, and silicon teams to define interfaces and validate behavior using event-based, cycle-accurate simulators, then diagnose correctness and performance issues. Add tests, tracing, profiling, and observability to ensure production-ready runtime capabilities.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Design and implement the low-level device runtime for OpenAI custom silicon.
  • •Build kernel-launch scheduling, command submission, queueing, dependency tracking, and completion handling.
  • •Manage device memory spaces, allocation, virtual-to-physical mappings, data movement, and lifetime across concurrent workloads.
  • •Implement synchronization primitives, events, barriers, streams, and ordering guarantees that are correct and efficient.
  • •Use event-based, cycle-accurate simulators to develop, validate, debug, and performance-tune runtime behavior before and after silicon availability.

Pay and Benefits

Salary: USD 266,000 - 445,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Strong low-level systems programming experience in C, C++, Rust, or comparable environments.
  • •Experience building runtimes, drivers, firmware, operating-system components, accelerator software, or adjacent infrastructure.
  • •Deep understanding of concurrency, synchronization, asynchronous execution, queues, events, and memory-ordering semantics.
  • •Experience with memory management including address spaces, DMA, caching, coherency, and hardware-software interfaces.
  • •Hands-on debugging across software abstractions, device interfaces, and hardware behavior.
Experience:Hardware-software co-design
Skills:DebuggingCollaborationQuantitative reasoningProblem-solving
Tech Stack:CC++Rust

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor